Skip to content
AI.info

jobs

Agent Post-Training, Frontier Evals and Environments Research

OpenAI's Agent Post-Training team develops the frontier agents used across Codex, ChatGPT, the API, and other products. This role focuses on building evaluation environments and training systems that guide progress toward safe, capable arti

Company
OpenAI
Location
San Francisco, United States
Status
Closed
Posted
2026-09-18T03:33:24.451+00:00

OpenAI's Agent Post-Training team works on the frontier agents behind Codex, ChatGPT and the API. This role builds reinforcement-learning environments and evaluation systems: creating ambitious RL environments, measuring model capabilities and behavior, developing methods to explore model behavior automatically, and improving the reliability and statistical quality of evaluations. It may include steering training runs and building continuous-evaluation systems. Candidates typically need strong foundations in machine learning, software engineering, systems or statistics, plus experience with LLMs, post-training, RLHF/RLAIF, model evaluations, graders, synthetic data, coding or tool-using agents. Hybrid, San Francisco.

Original job posting