HHiring Reality
← Together AI

Forward Deployed Engineer (Inference & Post-Training)

Together AI
Location
San Francisco
Posted
12 days ago
Department
Customer Success
Salary
$270,000 - $300,000
What they actually want (must-haves)
  • 5+ years in a technical role with a focus on inference systems or post-training workflows
  • Expert-level experience with inference engines (e.g., vLLM, TensorRT-LLM)
  • Deep knowledge of KV cache tuning, speculative decoding, tensor parallelism, and quantization techniques
  • Hands-on experience with fine-tuning and post-training pipelines (LoRA, SFT, DPO, RLHF, GRPO)
  • Strong Python skills and experience in production environments
  • Broad knowledge of state-of-the-art open-source models and model selection for customer use cases
What the job really is

As a Forward Deployed Engineer focused on Inference & Post-Training, you will work closely with strategic customers to optimize their AI models for production use. Your role involves hands-on technical support in areas such as inference engine optimization, performance tuning, and guiding customers through post-training workflows. You will also contribute to product feedback and help ensure successful platform adoption by aligning with customer needs from onboarding through deployment.

Benefits
  • Competitive compensation
  • Startup equity
  • Health insurance
  • Flexibility in terms of remote work
Things to weigh
  • No specific mention of career growth opportunities or training programs
  • The role requires deep technical expertise, which may limit accessibility for less experienced candidates
Job score4.2/5
Benefits2/5
Freshness5/5
Career value4/5
Role clarity5/5
Pay transparency5/5

Applying to Together AI?

See how your résumé matches this role — and tailor it from what actually gets interviews.