Tech Lead, Deployment & Operations — Custom InfrastructureActive$342K–$445K

The opportunity

OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models.

What you'll do

  • Lead a team responsible for deployment and operations of OpenAI’s custom: silicon and systems in data center environments

  • Own the path from hardware bring-up and validation through production: deployment, operational readiness, and sustained fleet support

  • Partner closely with silicon, systems, software, infrastructure, networking,: data center, supply chain, and external partner teams to ensure successful deployment at scale

  • Define deployment processes, operational playbooks, technical readiness: criteria, escalation paths, and reliability practices for new hardware platforms

  • Drive cross-functional execution across lab bring-up, rack/system: integration, data center deployment, fleet monitoring, debugging, and issue resolution

  • Stay hands-on technically through architecture reviews, deployment planning,: failure analysis, operational debugging, and critical system-level decision-making

What they're looking for

  • + years of engineering experience in hardware systems, infrastructure, data: center deployment, production operations, systems engineering, silicon bring-up, or related technical domains
  • Strong technical depth in one or more of: hardware deployment, data center operations, rack-scale systems, silicon bring-up, systems validation, fleet operations, reliability engineering, infrastructure automation, or hardware/software integration
  • Experience bringing complex hardware systems from development or validation into production environments
  • Experience working closely with silicon, systems, software, infrastructure, networking, or data center teams
  • Experience with deployment planning, operational readiness, incident: response, debugging, and root-cause analysis for production systems
  • Experience building tooling, automation, observability, or operational: processes that improve deployment quality and fleet reliability
  • Demonstrated ability to hire, develop, and lead senior technical talent
  • Ability to move fluidly between people leadership, technical strategy, and: hands-on operational problem solving
Tech Lead, Deployment & Operations — Custom Infrastructure at OpenAI | Role Match