Vai al contenuto
AI.info

jobs

Staff Software Engineer, Inference Cloud

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference serv

In inglese

Company
Cerebras
Location
Sunnyvale, CA
Status
Open
Posted
2024-07-12T03:15:09.67+00:00

Staff engineer on Cerebras's Inference Cloud Platform team, the cloud layer behind its inference service, which the posting says runs on the company's wafer-scale AI chips. The role owns major architectural areas: multi-region topology, service discovery, request routing, load balancing, caching, batching, admission control and quota management, plus reliability work on active-active failover and SLOs. It asks for 8+ years in software engineering, deep distributed-systems expertise, Go, C++ or Python, and high-QPS latency work. ML inference or model-serving experience is preferred.

Original job posting