jobs
Machine Learning Research Scientist, Evaluations
Scale works with the industry's leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward mod
In inglese
- Company
- Scale AI
- Location
- San Francisco, CA; Seattle, WA; New York, NY
- Status
- Open
- Posted
- 2026-08-26T18:51:42+00:00
Scale AI is hiring research scientists and engineers for the evaluation pod within its GenAI Research Organization. The work involves diagnosing how frontier language and multimodal models fail — capability gaps, reasoning errors, robustness and alignment — and building benchmarks and evaluation methods to measure capabilities. The posting asks for a PhD or master's in computer science or a related field, post-training experience with SFT, RLHF or reward modelling, and published conference research. Customer-facing experience is preferred. Roles are on-site in San Francisco, Seattle or New York.
Original job posting