HHiring Reality
← Together AI

AI Infrastructure Systems Engineer (Amsterdam & London)

Together AI
Location
Amsterdam
Posted
2 months ago
Department
Engineering
What they actually want (must-haves)
  • 3+ years building distributed systems, infrastructure platforms, or large-scale backend software
  • Strong software engineering skills in Python, Go, or Rust
  • Experience building platforms, automation systems, or developer infrastructure
  • Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies
  • Strong systems thinking with the ability to understand problems across hardware and software
  • A passion for solving complex infrastructure challenges through software
Nice to have
  • GPU infrastructure, CUDA, NCCL, NVLink/NVSwitch
  • InfiniBand or RoCE networking
  • Bare-metal provisioning and lifecycle management
  • Large-scale AI training or inference clusters
  • Hardware health monitoring and predictive failure detection
What the job really is

In this role at Together AI, you will design and build automation systems for managing a large fleet of GPUs used in AI model training and inference. Your responsibilities will include creating software for fleet automation, monitoring hardware health, and improving operational efficiency through automation. You will work closely with various teams to enhance the AI infrastructure and ensure high performance and reliability.

Things to weigh
  • Role involves building and operating large-scale GPU infrastructure, which may require specialized knowledge
  • No specific salary or benefits mentioned in the posting
  • Hybrid work model available, but remote work is limited to the UK
Job score2.4/5
Benefits1/5
Freshness1/5
Career value5/5
Role clarity5/5
Pay transparency0/5

Applying to Together AI?

See how your résumé matches this role — and tailor it from what actually gets interviews.