In this role, you will be responsible for building a Kubernetes-native control plane to manage a GPU inference fleet. Your tasks will include designing APIs for the inference team, automating self-healing processes, and ensuring the reliability of the provisioning pipeline. You will also collaborate with the inference/ML platform team to encode necessary cluster shapes and engineer the infrastructure code with a focus on quality and developer experience.
See how your résumé matches this role — and tailor it from what actually gets interviews.