The opportunity
SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.
What you'll do
Develop automation to deploy and manage on-premise Kubernetes clusters
Deploy and manage core infrastructure such as databases, monitoring and distributed storage
Closely collaborate with software engineers to create highly scalable, operable, and maintainable products
Engage in and improve the whole lifecycle of services -: from inception and design, through deployment, operation and refinement
Monitoring and alerting supporting systems to have high availability
Hands-on integration and troubleshooting across the entire Starlink stack
What they're looking for
- Identify areas for improvement and create innovative solutions that enable high system availability
- Bachelor’s degree in computer science, information systems/IT, or an: engineering discipline and 1+ years of professional experience in Site Reliability Engineering or DevOps; OR 3+ years of professional experience in Site Reliability Engineering or DevOps in lieu of a degree
- + years of professional experience with Linux operating systems
- Experience with Terraform, Ansible, or other infrastructure tools