The opportunity
SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.
What you'll do
Apply expertise to large-scale compute system development, including design: validation, integration, performance tuning, and optimization.
Act as the primary technical point of contact for compute infrastructure: definition, requirements, and delivery across all customer accounts.
Lead technical implementation, integration testing, provisioning, and go-live: activities for GPU/CPU clusters and supporting infrastructure.
Provide deep expertise on compute-specific elements including GPUs/CPUs,: high-performance networking, data center connectivity, and related systems.
Handle technical escalations, root-cause analysis, troubleshooting, and drive: resolution to ensure SOW and SLA compliance.
Represent compute systems in cross-functional trades, risk discussions, and: issue resolution with internal teams, management, and customers.
What they're looking for
- Interface with Hardware, Network, Facilities, Cloud Compute, SRE, and Program: Management teams to ensure successful end-to-end delivery.
- Participate in verification testing, performance characterization,: reliability assessments, and continuous optimization.
- Support expansion planning, capacity scaling, and new use cases across customer accounts.
- Ensure on-time deliverables, proactive risk mitigation, and overall mission success for all customers.