Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS)Active

The opportunity

MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance…

What you'll do

  • Work on our multi-tenant distributed storage systems, balancing long-term: strategic infrastructure goals with immediate engineering needs

  • Build for reliability, making services and infrastructure available,: resilient, fault-tolerant, and self-healing

  • Identify and configure key metrics to detect incidents and quantify service: health, availability, and performance

  • Participate in a 24/7 on-call rotation to resolve issues involving the storage infrastructure

  • Become an expert in infrastructure performance, helping us optimize from the: application level all the way to the kernel