The opportunity
OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models.
What you'll do
Lead a team responsible for deployment and operations of OpenAI’s custom: silicon and systems in data center environments
Own the path from hardware bring-up and validation through production: deployment, operational readiness, and sustained fleet support
Partner closely with silicon, systems, software, infrastructure, networking,: data center, supply chain, and external partner teams to ensure successful deployment at scale
Define deployment processes, operational playbooks, technical readiness: criteria, escalation paths, and reliability practices for new hardware platforms
Drive cross-functional execution across lab bring-up, rack/system: integration, data center deployment, fleet monitoring, debugging, and issue resolution
Stay hands-on technically through architecture reviews, deployment planning,: failure analysis, operational debugging, and critical system-level decision-making
What they're looking for
- + years of engineering experience in hardware systems, infrastructure, data: center deployment, production operations, systems engineering, silicon bring-up, or related technical domains
- Strong technical depth in one or more of: hardware deployment, data center operations, rack-scale systems, silicon bring-up, systems validation, fleet operations, reliability engineering, infrastructure automation, or hardware/software integration
- Experience bringing complex hardware systems from development or validation into production environments
- Experience working closely with silicon, systems, software, infrastructure, networking, or data center teams
- Experience with deployment planning, operational readiness, incident: response, debugging, and root-cause analysis for production systems
- Experience building tooling, automation, observability, or operational: processes that improve deployment quality and fleet reliability
- Demonstrated ability to hire, develop, and lead senior technical talent
- Ability to move fluidly between people leadership, technical strategy, and: hands-on operational problem solving