Staff Software Engineer, Foundation Model InferenceActive$190K

The opportunity

At Databricks, we are passionate about enabling data and AI teams to solve the world's toughest problems — from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world's best data and…

What you'll do

  • Build LLM infrastructure powering large-scale inference workloads for: customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS, Llama)

  • Shape the direction of the FMAPI product: from roadmap to execution — by leveraging deep customer empathy and direct engagement with enterprise users and model providers

  • Improve reliability, latency, and efficiency of distributed AI workloads

  • Collaborate with platform, infra, and ML teams to deliver seamless end-to-end experiences

  • Shape how developers and data scientists build and interact with AI on Databricks

  • + years of experience in backend or infrastructure engineering

What they're looking for

  • Experience with distributed systems, scalable APIs, or cloud-native infrastructure
  • Strong product and ownership mindset, with a focus on shipping user-facing value
  • Experience with real-time serving, ML infrastructure, or GPU orchestration
  • Familiarity with service-oriented architecture, deployment pipelines, and system observability