jobs
Staff Software Engineer- Foundation Model Inference
P-1930 At Databricks, we are passionate about enabling data and AI teams to solve the world's toughest problems — from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by
- Company
- Databricks
- Location
- San Francisco, California
- Status
- Open
- Posted
- 2026-07-24T16:49:01+00:00
Databricks is hiring a staff software engineer for its Foundation Model Inference team, which builds the infrastructure behind the company's generative AI products. The work covers large-scale LLM inference for partner models such as OpenAI, Anthropic and Gemini, and self-hosted models including Llama and Qwen, plus reliability, latency and efficiency of distributed AI workloads. The posting asks for 8+ years in backend or infrastructure engineering, distributed systems or cloud-native infrastructure experience, and real-time serving or GPU orchestration. It gives preference to candidates with SageMaker, Vertex AI, Azure ML exposure or contributions to vLLM, SGLang, Ray, PyTorch or MLflow.
Original job posting