The ML Serving Engineer will design and implement high-performance ML serving architectures for game applications, ensuring stable and efficient operation of AI models in production environments. Responsibilities include building real-time serving infrastructure for large language models (LLMs), optimizing performance, and establishing CI/CD pipelines for continuous deployment and monitoring.
See how your résumé matches this role — and tailor it from what actually gets interviews.