The opportunity
SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.
What you'll do
Own the architecture, lifecycle, and day-to-day operation of enterprise: compute, virtualization, and storage platforms that support SpaceX engineering, production, and mission workloads.
Manage and continuously develop a team of highly capable engineers and: administrators who design, build, and operate hypervisor clusters, purpose-built servers, and high-performance storage systems (NAS, SAN, and object storage).
Lead capacity planning, performance engineering, and standards for compute: and storage so hardware and software platforms stay ahead of demand from SpaceX engineering teams.
Design, implement, and report on high-availability, backup, replication, and: disaster recovery strategies for virtual machines, datastores, and storage arrays.
Mentor systems engineers in designing, configuring, and supporting: purpose-built servers and virtualization hosts, with emphasis on performance-optimized, resilient designs.
Establish hardware and platform standards for servers, hypervisors, and: storage; submit and track orders; and work with server and storage vendors to negotiate pricing and ensure availability.
What they're looking for
- Supervise installation, configuration, firmware/lifecycle management, and: decommission of servers, hypervisor hosts, and storage systems.
- Drive automation of provisioning, patching, and operational tasks across: compute and storage (templates, infrastructure as code, and scripting) so the team can scale without linear headcount growth.
- Partner with network, facilities, security, and application teams on: connectivity, power/cooling constraints, access controls, and workload placement — without owning data center mechanical/electrical systems.
- Schedule and provide off-peak or weekend support when necessary to perform: high-risk or planned downtime of SpaceX compute, virtualization, and storage systems for upgrades and maintenance.