The opportunity
The Product & Platform teams at OpenAI are responsible for delivering the company’s most impactful offerings—such as ChatGPT, our API platform, and new enterprise capabilities—to a global and diverse customer base. These systems must perform at scale and deliver exceptional…
What you'll do
Own cross-functional programs for Chat capacity forecasting, allocation,: headroom planning, and constrained-capacity operations.
Build durable intake, prioritization, and decision mechanisms that connect: product demand and model requirements to available serving capacity.
Partner with product, research, inference, fleet, and capacity teams to: develop scenarios, surface tradeoffs, and drive timely allocation decisions.
Lead model deployment readiness and rollout planning, including: serving-capacity allocation, launch sequencing, validation, and operational handoffs.
Establish clear readiness gates, risk reviews, rollback criteria, and escalation paths for model deployments.
Drive launch coordination through deployment and post-launch learning,: turning recurring gaps and manual work into scalable tooling and operating practices.
What they're looking for
- Define and operationalize metrics for forecast accuracy, capacity utilization: and headroom, deployment velocity, reliability, latency, quality, and user impact.
- Create concise, decision-ready communications that make dependencies, risks,: capacity constraints, and launch choices clear to technical and product leaders.
- Have led complex technical programs in infrastructure, distributed systems,: capacity planning, model serving, or large-scale deployment environments.
- Can reason credibly about demand, supply, headroom, reliability, latency, and: quality tradeoffs, and translate them into executable plans.