In this role, you will develop software systems that automate the lifecycle management of hardware infrastructure, transforming physical servers into operational AI clusters. You will create APIs for self-service provisioning, implement state machines for hardware management, and ensure reliability through automated self-healing processes. The position emphasizes a product mindset, requiring you to build, own, and support the software you develop.
See how your résumé matches this role — and tailor it from what actually gets interviews.