Vai al contenuto
AI.info

jobs

Agent Post-Training, Computer Use Research

About the TeamThe Agent Post-Training team creates the frontier agents OpenAI ships to the world. We are training the models behind our agents in Codex, ChatGPT, the API, and other frontier products: persistent, proactive intelligence that

In inglese

Company
OpenAI
Location
San Francisco
Status
Open
Posted
2026-06-26T17:42:39.603+00:00

OpenAI's Agent Post-Training team, on its computer-use effort, hires a researcher to teach models to operate browsers and desktops. The work covers RL and post-training pipelines, graders, reward signals, evals, environments and data, run through major training runs and into Codex and ChatGPT. The posting asks for technical fundamentals in ML or systems, hands-on experience with LLMs, RLHF/RLAIF or agent training, and comfort moving from vague behavioral failures to concrete experiments. It is unusual for spanning research taste, product impact and load-bearing engineering.

Original job posting