The opportunity
Elastic, the Search AI Company, enables everyone to find the answers they need in real time, using all their data, at scale — unleashing the potential of businesses and people. The Elastic Search AI Platform, used by more than 50% of the Fortune 500, brings together the…
What you'll do
Owning end-to-end delivery of moderate-to-high complexity projects on the: team’s roadmap, with minimal day-to-day direction.
Operating and hardening shared Elastic Cloud infrastructure (ECH, ECE, and: ECK) as Infrastructure as Code – writing and reviewing the Terraform, Python, and Go that other engineers depend on.
Carrying a 24/7 on-call rotation: responding to incidents, driving them to resolution, and writing clear RCAs/postmortems that lead to lasting fixes rather than repeat pages.
Reviewing others’ code and designs, and being a trusted second set of eyes on: production changes to critical infrastructure.
Mentoring less experienced engineers, and proactively raising risks, ideas,: and improvements in team discussions.
Improving runbooks, documentation, and operational processes so the on-call load gets lighter over time.
What they're looking for
- + years of SRE, platform engineering, or infrastructure engineering experience
- Proficiency with Terraform; comfortable owning large, multi-workspace configurations in a team setting
- Strong software engineering fundamentals in Python; comfort with Go is a plus.
- Deep Linux systems knowledge and experience operating containerized workloads in production.