Senior Site Reliability Engineer (Observability & Analytics) – Platform InfraPosted today$46K–$183K

The opportunity

Elastic, the Search AI Company, enables everyone to find the answers they need in real time, using all their data, at scale — unleashing the potential of businesses and people. The Elastic Search AI Platform, used by more than 50% of the Fortune 500, brings together the…

What you'll do

  • Owning end-to-end delivery of moderate-to-high complexity projects on the: team’s roadmap, with minimal day-to-day direction.

  • Operating and hardening shared Elastic Cloud infrastructure (ECH, ECE, and: ECK) as Infrastructure as Code – writing and reviewing the Terraform, Python, and Go that other engineers depend on.

  • Carrying a 24/7 on-call rotation: responding to incidents, driving them to resolution, and writing clear RCAs/postmortems that lead to lasting fixes rather than repeat pages.

  • Reviewing others’ code and designs, and being a trusted second set of eyes on: production changes to critical infrastructure.

  • Mentoring less experienced engineers, and proactively raising risks, ideas,: and improvements in team discussions.

  • Improving runbooks, documentation, and operational processes so the on-call load gets lighter over time.

What they're looking for

  • + years of SRE, platform engineering, or infrastructure engineering experience
  • Proficiency with Terraform; comfortable owning large, multi-workspace configurations in a team setting
  • Strong software engineering fundamentals in Python; comfort with Go is a plus.
  • Deep Linux systems knowledge and experience operating containerized workloads in production.