jobs
Senior AI Inference Engineer - Model Optimization & Deployment
The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence. As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficien
- Company
- Zoox
- Location
- Foster City, CA
- Status
- Open
- Posted
- 2026-04-11T00:04:50.918+00:00
Zoox's Perception team is hiring a senior engineer to deploy large models onto the on-vehicle stack. The work involves compressing, accelerating and optimizing models such as LLMs, VLMs or foundation models for power- and thermal-constrained vehicle SOCs, writing custom CUDA kernels, and building concurrent inference code for real-time, deterministic execution on edge devices. The posting seeks hands-on experience with model compression and deployment for constrained hardware. What stands out is the focus on on-vehicle inference and the combination of kernel-level work with foundation-model deployment.
Original job posting