Senior Site Reliability Engineer IActive

The opportunity

At Braze, we have found our people. We’re a genuinely approachable, exceptionally kind, and intensely passionate crew.

What you'll do

  • Partner with Braze’s engineering teams on:

  • Architecting products to effectively utilize infrastructure platforms in a scalable, reliable manner

  • Debugging reliability and scalability issues across all stack layers,: including the products built using our infrastructure platforms

  • Make monitoring and alerting alerts on symptoms and not on outages

  • Ensure that Braze meets our strict enterprise-grade SLAs with customers

  • Develop Braze’s internal platform infrastructure:

What they're looking for

  • Create Infrastructure as code using Chef, Terraform, and Kubernetes
  • Develop deployment pipelines for applications in multiple languages using Docker, Kubernetes, etc
  • Provide centralized/common tooling, services, and automation frameworks that: are critical for scaling operations, capacity management, reducing operational pain, and improving the day-to-day workflow of Braze’s engineering teams
  • Manage incidents: