HHiring Reality
← Saviynt

AI Platform Engineer, Training and Inference

Saviynt
Location
Milpitas, California
Posted
2 months ago
Department
Platform Upgrade
What they actually want (must-haves)
  • Experience in ML engineering with time in an ML platform or MLOps role
  • Production Ray depth: Ray Train, Serve, Core, and Data
  • Hands-on with LLM serving engines: vLLM, SGLang, or NVIDIA Triton
  • Knowledge of distributed training: DDP, FSDP, NCCL collectives, gradient checkpointing, and mixed precision
  • Familiarity with model lifecycle operations: MLflow registry, shadow/A/B/canary patterns, and auto-rollback
  • Experience with vector databases: Pgvector or Qdrant
Nice to have
  • Quantization: INT8/INT4/FP8 post-training quantization
What the job really is

As an AI Platform Engineer focused on Training and Inference at Saviynt, you will be responsible for managing and optimizing the Ray ecosystem for distributed training and inference of AI models. Your day-to-day tasks will involve operating the full model promotion lifecycle, building RL training infrastructure, and ensuring the performance of LLM inference mesh. You will also work on integrating retrieval-augmented generation into the inference process, contributing to the development of AI solutions that enhance Saviynt's identity products.

Benefits
  • Competitive total rewards package
  • Eligibility to participate in a discretionary bonus plan
  • Opportunities for growth and advancement in your career
Things to weigh
  • No specific salary range listed
  • Role involves managing complex distributed systems, which may require a steep learning curve
  • Emphasis on security and compliance may add additional responsibilities
Job score2.6/5
Benefits2/5
Freshness1/5
Career value5/5
Role clarity5/5
Pay transparency0/5

Applying to Saviynt?

See how your résumé matches this role — and tailor it from what actually gets interviews.