jobs
Cluster Operations Software Engineer
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference serv
- Company
- Cerebras
- Location
- Sunnyvale, CA; Toronto, CAN; Bengaluru, IND
- Status
- Open
- Posted
- 2026-08-13T16:10:31.749+00:00
Cerebras is hiring an engineer to run the machine-learning compute clusters built around its Wafer-Scale Engine. The job covers deploying and debugging Docker-based services, writing Python and Go tooling for monitoring, automation, dashboards and reliability, managing cluster health and capacity, and taking part in a 24/7 on-call rotation. Cerebras asks for six to eight years with complex compute infrastructure, distributed systems, Linux and Kubernetes. The work spans three sites and interacts directly with the company's own AI hardware.
Original job posting