HHiring Reality
← Together AI

Junior/Senior or Staff Software Engineer, Inference / Compute Infrastructure Engineering

Together AI
Location
India
Posted
1 month ago
Department
Engineering
What they actually want (must-haves)
  • Strong software engineering background in Go, Python, Rust, or similar
  • Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent
  • Experience building software control planes or orchestration systems that model state and reconcile it over time
  • Experience with event-driven systems
  • A product mindset with experience building internal platforms or APIs consumed by other engineering teams
Nice to have
  • Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals
  • Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE)
  • Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization
  • Systems programming in Rust or Go
What the job really is

In this role, you will be responsible for building a Kubernetes-native control plane to manage a GPU inference fleet. Your tasks will include designing APIs for the inference team, automating self-healing processes, and ensuring the reliability of the provisioning pipeline. You will also collaborate with the inference/ML platform team to encode necessary cluster shapes and engineer the infrastructure code with a focus on quality and developer experience.

Things to weigh
  • Remote position in India
  • No specific salary or benefits mentioned
  • Focus on operational and runtime complexity may require deep technical understanding
  • Role involves both development and production support responsibilities
Job score2.8/5
Benefits1/5
Freshness4/5
Career value4/5
Role clarity5/5
Pay transparency0/5

Applying to Together AI?

See how your résumé matches this role — and tailor it from what actually gets interviews.