Skip to content
AI.info

jobs

Software Engineer, GPU Infrastructure - HPC

About the teamThe Fleet team at OpenAI supports the computing environment that powers our cutting-edge research and product development. We oversee large-scale systems that span data centers, GPUs, networking, and more, ensuring high availa

Company
OpenAI
Location
San Francisco; New York City
Status
Open
Posted
2026-02-05T19:14:42.993+00:00

OpenAI's Fleet team is hiring a software engineer for its High Performance Computing group, which keeps the company's GPU compute fleet reliable and available. The work covers automation for provisioning and managing server fleets, tooling to monitor server health, performance and lifecycle events, and hunting down performance bottlenecks alongside clusters and networking teams. The posting asks for experience running large-scale server environments, Python or Go, strong Linux, networking and server-hardware knowledge, and comfort with SQL, PromQL or Pandas. Hardware expertise is explicitly not required.

Original job posting