jobs
Forward Deployed Engineer (Inference & Post-Training)
About the role As a Forward Deployed Engineer (FDE) focused on Inference & Post-Training, you will be a hands-on technical partner to our most strategic customers — production AI teams looking to leverage high quality models and do inferenc
- Company
- Together AI
- Location
- San Francisco
- Status
- Open
- Posted
- 2026-05-07T20:52:41+00:00
Together AI is hiring a Forward Deployed Engineer for inference and post-training. The posting says the role works alongside solutions architects rather than replacing them. Work includes optimizing inference engines, tuning configurations, and guiding fine-tuning or RL pipelines for strategic accounts. Responsibilities cover KV cache tuning, speculative decoding, quantization, tensor parallelism, and hands-on LoRA, SFT, DPO, RLHF, and GRPO work. The posting asks for 5+ years in a technical role, expert use of vLLM, TensorRT-LLM, or SGLang, Python skills, and broad open-model knowledge.
Original job posting