Skip to content
AI.info

jobs

Senior AI Inference Engineer - Model Optimization & Deployment

The Perception team is pioneering the development of a multi-modality foundation model to drive the next generation of autonomous system intelligence. As a Model Optimization & Deployment Engineer, you will focus on bringing highly efficien

Company
Zoox
Location
Foster City, CA
Status
Open
Posted
2026-04-11T00:04:50.918+00:00

Zoox's Perception team is hiring a senior engineer to deploy large models onto the on-vehicle stack. The work involves compressing, accelerating and optimizing models such as LLMs, VLMs or foundation models for power- and thermal-constrained vehicle SOCs, writing custom CUDA kernels, and building concurrent inference code for real-time, deterministic execution on edge devices. The posting seeks hands-on experience with model compression and deployment for constrained hardware. What stands out is the focus on on-vehicle inference and the combination of kernel-level work with foundation-model deployment.

Original job posting