Skip to content
AI.info

jobs

Research Scientist, Frontier Risk Evaluations

Scale Labs, Research Scientist — Frontier Risk Evaluations As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building

Company
Scale AI
Location
San Francisco, CA; New York, NY
Status
Open
Posted
2026-03-25T21:15:14+00:00

Scale Labs, Scale AI's policy research arm, seeks a research scientist to design evaluations measuring risks from frontier AI systems. Work includes building harnesses and datasets that test models and agents for dangerous capabilities such as security vulnerability exploitation and CBRN uplift, scoping evaluations with government agencies and other labs, and publishing methodologies and technical reports for policymakers. The posting asks for published ML research, at least three years working on complex ML problems, and comfort building ML pipelines and evaluation harnesses. Interviews test practical prototyping and debugging, not LeetCode.

Original job posting