Skip to content
AI.info

jobs

Machine Learning Research Scientist, Evaluations

Scale works with the industry's leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward mod

Company
Scale AI
Location
San Francisco, CA; Seattle, WA; New York, NY
Status
Open
Posted
2026-08-26T18:51:42+00:00

Scale AI is hiring research scientists and engineers for the evaluation pod within its GenAI Research Organization. The work involves diagnosing how frontier language and multimodal models fail — capability gaps, reasoning errors, robustness and alignment — and building benchmarks and evaluation methods to measure capabilities. The posting asks for a PhD or master's in computer science or a related field, post-training experience with SFT, RLHF or reward modelling, and published conference research. Customer-facing experience is preferred. Roles are on-site in San Francisco, Seattle or New York.

Original job posting