Software Engineer, Inference - Multi ModalActive$295K–$555K

San FranciscoTechnology

The opportunity

About the Team OpenAI’s Inference team powers the deployment of our most advanced models - including our GPT models, 4o Image Generation, and Whisper - across a variety of platforms. Our work ensures these models are available, performant, and scalable in production, and we…

What they're looking for

  • Have worked with GPU-based ML workloads and understand the performance: dynamics of large models, especially with complex data like images or audio.
  • Enjoy experimental, fast-evolving work and collaborating closely with research.
  • Are comfortable dealing with systems that span networking, distributed: compute, and high-throughput data handling.
  • Have familiarity with inference tooling like vLLM, TensorRT-LLM, or custom model parallel systems.