The opportunity
About the Team OpenAI’s Inference team powers the deployment of our most advanced models - including our GPT models, 4o Image Generation, and Whisper - across a variety of platforms. Our work ensures these models are available, performant, and scalable in production, and we…
What they're looking for
- Have worked with GPU-based ML workloads and understand the performance: dynamics of large models, especially with complex data like images or audio.
- Enjoy experimental, fast-evolving work and collaborating closely with research.
- Are comfortable dealing with systems that span networking, distributed: compute, and high-throughput data handling.
- Have familiarity with inference tooling like vLLM, TensorRT-LLM, or custom model parallel systems.