jobs
Researcher, Interpretability
About the TeamThe Interpretability team studies internal representations of deep learning models. We are interested in using representations to understand model behavior, and in engineering models to have more understandable representations
- Company
- OpenAI
- Location
- San Francisco
- Status
- Open
- Posted
- 2025-06-15T20:14:27.535+00:00
OpenAI's Interpretability team is hiring a researcher to study internal representations of deep learning models, aiming to understand model behavior, make representations more interpretable, and apply that to AI safety. The work includes publishing research on techniques for understanding deep networks, engineering infrastructure to study model internals at scale, collaborating across teams, and steering research toward usefulness or long-term scalability. The posting asks for a Ph.D. or research experience in computer science or machine learning, 2+ years of research engineering experience, and proficiency in Python.
Original job posting