The opportunity
OpenAI is helping build the infrastructure that powers the next generation of artificial intelligence. Through Stargate, we are developing and operating large-scale AI compute campuses that require world-class execution across data center design, construction, commissioning, and operations.
What you'll do
Lead day-to-day operations of mission-critical facility infrastructure across AI compute campuses.
Own operational readiness activities supporting new campus deployments and infrastructure expansion.
Partner with commissioning teams to transition facilities from construction: and startup into steady-state operations.
Develop, implement, and continuously improve operating procedures, maintenance programs, and response plans.
Lead infrastructure incident response efforts and coordinate recovery activities during critical events.
Drive root cause analysis investigations and corrective action programs to: improve reliability and operational performance.
What they're looking for
- Manage vendors, contractors, and service providers supporting facility operations.
- Partner with hardware deployment, networking, and engineering teams to: coordinate infrastructure changes and maintenance activities.
- Monitor facility performance, operational risk, and capacity utilization across critical systems.
- Support staffing, training, and development of facilities operations personnel.