jobs
Staff Software Engineer, Inference Cloud
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference serv
In inglese
- Company
- Cerebras
- Location
- Sunnyvale, CA
- Status
- Open
- Posted
- 2024-07-12T03:15:09.67+00:00
Staff engineer on Cerebras's Inference Cloud Platform team, the cloud layer behind its inference service, which the posting says runs on the company's wafer-scale AI chips. The role owns major architectural areas: multi-region topology, service discovery, request routing, load balancing, caching, batching, admission control and quota management, plus reliability work on active-active failover and SLOs. It asks for 8+ years in software engineering, deep distributed-systems expertise, Go, C++ or Python, and high-QPS latency work. ML inference or model-serving experience is preferred.
Original job posting