jobs
Software Engineer, Data Infrastructure - Research
About the TeamThe Workload team is responsible for designing and running OpenAI’s LLM training and inference infrastructure that powers frontier models at massive scale. Our systems unify how researchers train and serve models, abstracting
In inglese
- Company
- OpenAI
- Location
- San Francisco
- Status
- Open
- Posted
- 2025-09-18T23:14:35.252+00:00
OpenAI's Workload team is hiring an engineer to build the dataset infrastructure behind its next-generation training stack: standardized dataset APIs, including for multimodal data too large to fit in memory, plus pipelines for loading data across thousands of GPUs. The work covers scale validation, debugging stragglers and other bottlenecks in distributed loading, reproducibility safeguards, and inspection tools. The posting asks for strong engineering fundamentals in distributed systems or data pipelines, comfort building APIs and abstractions, and GPU-scale experience as a bonus.
Original job posting