Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)Active$144K

The opportunity

MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance…

What you'll do

  • Have 6+ years of experience working on software development and operating distributed systems

  • Proficiency in Python, Go, or a similar language

  • Have operated or supported stateful storage or database systems at scale, and: are comfortable with durability, consistency, and recovery trade-offs.

  • Possess a customer-focused mindset

  • Value efficiency in processes and operations

  • Prefer automation over manual processes. We are a small team of software: engineers with a strong bias towards software solutions to avoid toil

What they're looking for

  • Experience using and extending containerization technologies, particularly: Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market
  • Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure
  • Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing)
  • Work on our multi-tenant distributed storage systems, balancing long-term: strategic infrastructure goals with immediate engineering needs