Skip to content
AI.info

jobs

AI Infrastructure Systems Engineer

About the Role At Together AI, you’ll build and operate one of the world’s largest GPU fleets used for frontier model training and inference. This isn’t a traditional infrastructure role—we’re looking for engineers who love building systems

Company
Together AI
Location
San Francisco
Status
Open
Posted
2026-05-15T03:04:22+00:00

Together AI is hiring a systems engineer to build and run automation for one of the largest GPU fleets, used for frontier model training and inference. The work covers software that provisions, validates, upgrades and repairs GPU clusters, plus agents for incident triage and monitoring of hardware, firmware, networking, storage and thermals. The posting asks for 3+ years of distributed systems or large-scale backend work, Python, Go or Rust, Linux and Kubernetes or Terraform, and an automation-first approach. It is on-site in San Francisco.

Original job posting