Site Reliability Engineer (Top Secret Clearance)Active$195K

The opportunity

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.

What you'll do

  • Develop automation to deploy and manage compute resources both on-premises and in the cloud

  • Build, maintain, and scale on-premises hardware systems designed to host: GPU-accelerated machine learning workloads

  • Deploy and manage core infrastructure such as databases, monitoring and storage

  • Closely collaborate with software engineers to create highly scalable, operable and maintainable products

  • Engage in and improve the whole lifecycle of services -: from inception and design, through deployment, operation and refinement

  • Bachelor’s degree in computer science, information systems, or an engineering: discipline; OR 2+ years of professional experience in software, DevOps, or site reliability engineering in lieu of a degree

What they're looking for

  • + year of experience with Kubernetes
  • + year of experience with Linux operating systems
  • Experience in Bash, Python, and/or other scripting languages
  • Experience building, maintaining, and scaling on-premises and/or cloud systems