Senior Site Reliability Engineer, FabricActive$127K

The opportunity

Platform Engineering sits within SRE and builds the core infrastructure powering MongoDB’s broader engineering organization. Our teams own everything from our multi-cloud Kubernetes foundation to automated deployment pipelines and global observability platforms.

What you'll do

  • Be a US Citizen, as the position supports our FedRAMP authorized environment

  • Have 5+ years of experience working on software and operating distributed: systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles

  • Be intimately familiar with modern cloud-based infrastructure and the network: design primitives of at least one of AWS, Azure, or GCP, e.g. VPCs, subnetting, routing, VPNs, peering, private link / private service connect, and CDNs

  • Possess a customer-focused mindset, driving improvements that benefit end-users

  • Value efficiency in processes and operations, and display a strong preference: for automation over manual processes (“allergic to ops work”)

  • Have a strong knowledge of service mesh and load-balancing concepts, and be: eager to implement these in a multi-cloud environment

What they're looking for

  • Contribute to a significant re-architecture of our multi-cloud: globally-connected network, bringing your experience to design discussions and owning meaningful parts of the work
  • Collaborate with service-owning teams to provide internal support, addressing: technical issues and offering guidance on best practices for service-to-service connectivity
  • Participate in a 24/7 on-call rotation to swiftly resolve issues related to: network architecture and service-to-service connectivity, ensuring minimal disruption and high availability