The opportunity
We’re hiring a Developer Productivity engineer to support OpenAI’s Inference Runtime teams. These teams own the systems responsible for serving models reliably, efficiently, and safely across Codex, ChatGPT, API, and internal research workloads.
What you'll do
Improve systems that ensure inference engine releases are correct,: performant, and regression-free by evolving tooling and infrastructure for deploy gate validation
Bring rigor to release, validation, branching, and deployment processes across the inference stack
Improve canary, async, and large-scale validation workflows for inference systems
Harden CI, testing, and validation infrastructure so failures are actionable and trustworthy
Reduce noisy or flaky failures caused by infrastructure instability, GPU: scheduling, or test environment issues
Build automation for failure triage, ownership detection, debugging, and escalation
What they're looking for
- Partner closely with inference teams, research developer productivity, engine: acceleration, and infrastructure teams to improve release quality and rollout safety
- Reduce developer friction in testing, debugging, and release workflows so: engineers can move faster with confidence
- You have strong experience with CI/CD systems, testing infrastructure,: release tooling, developer productivity, or large-scale build and validation systems
- You are excited by high-impact infrastructure where small regressions in: correctness, latency, or reliability meaningfully affect production systems