Skip to content
AI.info

jobs

Technical Program Manager, Model Performance

ABOUT BASETENBaseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless dev

Company
Baseten
Location
San Francisco; Remote; Toronto; New York; Montreal; Seattle
Status
Open
Posted
2026-08-29T17:04:48.113+00:00

Baseten is hiring its first technical program manager for the Model Performance organization, which builds the algorithms behind its inference stack. The role is zero-to-one: the hire designs planning structures, operating cadences and status reporting, runs the project portfolio, and coordinates model release and optimization programs, including day-zero launches, across performance engineering, infrastructure and release teams. Requirements include program-managing inference optimization work and familiarity with serving engines such as vLLM, TensorRT-LLM, SGLang or NVIDIA Dynamo, plus influencing without authority. Hybrid; mid level.

Original job posting