Vai al contenuto
AI.info

jobs

Staff Software Engineer- Foundation Model Inference

P-1930 At Databricks, we are passionate about enabling data and AI teams to solve the world's toughest problems — from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by

In inglese

Company
Databricks
Location
San Francisco, California
Status
Open
Posted
2026-07-24T16:49:01+00:00

Databricks is hiring a staff software engineer for its Foundation Model Inference team, which builds the infrastructure behind the company's generative AI products. The work covers large-scale LLM inference for partner models such as OpenAI, Anthropic and Gemini, and self-hosted models including Llama and Qwen, plus reliability, latency and efficiency of distributed AI workloads. The posting asks for 8+ years in backend or infrastructure engineering, distributed systems or cloud-native infrastructure experience, and real-time serving or GPU orchestration. It gives preference to candidates with SageMaker, Vertex AI, Azure ML exposure or contributions to vLLM, SGLang, Ray, PyTorch or MLflow.

Original job posting