jobs
Research Engineer, Post-Training Inference
About the role The Model Shaping team at Together AI works on products and research focused on tailoring open foundation models to downstream applications. We build services that enable machine learning developers to choose the best models
- Company
- Together AI
- Location
- San Francisco
- Status
- Open
- Posted
- 2026-07-06T18:21:40+00:00
Together AI's Model Shaping team tailors open foundation models to downstream applications. The research engineer will build the platform that lets developers customize open-source models with their own data, wiring fine-tuning, reinforcement learning and evaluation services into the inference stack, optimizing engines for RL workloads, and joining an on-call rotation. The posting asks for 2+ years deploying ML services in production, hands-on work with SGLang, vLLM or TensorRT-LLM, familiarity with current fine-tuning methods, and Python or Go. Multi-LoRA serving, low-precision kernels, Triton/CUDA work and Kubernetes experience stand out.
Original job posting