Staff Software Engineer, Foundation Model InferenceActive$190K
The opportunity
At Databricks, we are passionate about enabling data and AI teams to solve the world's toughest problems — from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world's best data and…
What you'll do
Build LLM infrastructure powering large-scale inference workloads for: customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS, Llama)
Shape the direction of the FMAPI product: from roadmap to execution — by leveraging deep customer empathy and direct engagement with enterprise users and model providers
Improve reliability, latency, and efficiency of distributed AI workloads
Collaborate with platform, infra, and ML teams to deliver seamless end-to-end experiences
Shape how developers and data scientists build and interact with AI on Databricks
+ years of experience in backend or infrastructure engineering
What they're looking for
- Experience with distributed systems, scalable APIs, or cloud-native infrastructure
- Strong product and ownership mindset, with a focus on shipping user-facing value
- Experience with real-time serving, ML infrastructure, or GPU orchestration
- Familiarity with service-oriented architecture, deployment pipelines, and system observability