HHiring Reality
← Together AI

AI infrastructure System Engineer Bangalore

Together AI
Location
Bangalore, India
Posted
2 months ago
Department
Engineering
What they actually want (must-haves)
  • 3+ years building distributed systems, infrastructure platforms, or large-scale backend software
  • Strong software engineering skills in Python, Go, or Rust
  • Experience building platforms, automation systems, or developer infrastructure
  • Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies
  • Strong systems thinking with the ability to understand problems across hardware and software
  • A passion for solving complex infrastructure challenges through software
Nice to have
  • GPU infrastructure, CUDA, NCCL, NVLink/NVSwitch
  • InfiniBand or RoCE networking
  • Bare-metal provisioning and lifecycle management
  • Large-scale AI training or inference clusters
  • Hardware health monitoring and predictive failure detection
What the job really is

In this role at Together AI, you will design and build automation systems for managing a large fleet of GPUs used in AI model training and inference. Your responsibilities will include developing software for fleet automation, monitoring hardware health, and improving operational efficiency through automation, all while collaborating with various teams to enhance AI infrastructure.

Things to weigh
  • No salary listed
  • Role emphasizes automation and software development over traditional infrastructure tasks
  • Focus on large-scale systems may require a strong problem-solving mindset
Job score2.4/5
Benefits1/5
Freshness1/5
Career value5/5
Role clarity5/5
Pay transparency0/5

Applying to Together AI?

See how your résumé matches this role — and tailor it from what actually gets interviews.