Software Engineer, Inference – AMD GPU EnablementActive$295K–$555K

San FranciscoTechnology

The opportunity

About the Team Our Inference team brings OpenAI’s most capable research and technology to the world through our products. We empower consumers, enterprises and developers alike to use and access our state-of-the-art AI models, allowing them to do things that they’ve never been able to before.

What they're looking for

  • Have experience writing or porting GPU kernels using HIP, CUDA, or Triton,: and care deeply about low-level performance.
  • Are familiar with communication libraries like NCCL/RCCL and understand their: role in high-throughput model serving.
  • Have worked on distributed inference systems and are comfortable scaling models across fleets of accelerators.
  • Enjoy solving end-to-end performance challenges across hardware, system libraries, and orchestration layers.