Skip to content
AI.info

jobs

Sr./Staff TPM - Inference Capacity

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference serv

Company
Cerebras
Location
Sunnyvale, CA; Remote; Toronto, CAN
Status
Open
Posted
2026-06-29T20:50:26.407+00:00

Cerebras is hiring a Senior or Staff technical program manager to own capacity planning and fleet strategy for its Inference Service organization. The work covers 6/12/26-week rolling capacity models across clusters, allocation and model placement, datacenter bring-up, weekly utilization reporting, incident postmortems, and adoption of an in-house planning tool. The posting asks for 5+ years of TPM or product operations experience in cloud infrastructure or large-scale ML serving, inference-stack familiarity, and SQL, Grafana, and Python or Flux. Distinctive: the role is executive-visible and spans engineering, product, SRE, and operations.

Original job posting