jobs
Member of Technical Staff - Research, Inference
About Us:AI needs a new infrastructure layer. We're building it at Modal.Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt t
In inglese
- Company
- Modal
- Location
- New York; San Francisco
- Status
- Open
- Posted
- 2026-07-05T19:32:02.659+00:00
Modal builds infrastructure for AI workloads and is hiring a research member of technical staff for inference. The person owns inference research bets end to end: speculative decoding, disaggregated prefill and decode, quantization to FP8 and INT4, KV-cache and memory management, and autoscaling for spiky serverless traffic. They train speculators against production traffic, work with customers alongside forward-deployed engineers, and maintain collaborations with outside labs. The posting asks for an LLM inference research or systems background, fluency across the serving stack, and in-person work in New York or San Francisco.
Original job posting