Skip to content
AI.info

jobs

Forward Deployed Engineer (Inference & Post-Training)

About the role As a Forward Deployed Engineer (FDE) focused on Inference & Post-Training, you will be a hands-on technical partner to our most strategic customers — production AI teams looking to leverage high quality models and do inferenc

Company
Together AI
Location
San Francisco
Status
Open
Posted
2026-05-07T20:52:41+00:00

Together AI is hiring a Forward Deployed Engineer for inference and post-training. The posting says the role works alongside solutions architects rather than replacing them. Work includes optimizing inference engines, tuning configurations, and guiding fine-tuning or RL pipelines for strategic accounts. Responsibilities cover KV cache tuning, speculative decoding, quantization, tensor parallelism, and hands-on LoRA, SFT, DPO, RLHF, and GRPO work. The posting asks for 5+ years in a technical role, expert use of vLLM, TensorRT-LLM, or SGLang, Python skills, and broad open-model knowledge.

Original job posting