The opportunity
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences…
What you'll do
Manage the ticket queue prioritizing and resolving requests while identifying: recurring categories to automate or deflect
Participate in a rotating on-call and incident response schedule, including: weekends, troubleshooting and documenting issues in real time
Build and maintain monitoring dashboards (Tableau, Superset, Grafana) that: track service health, availability, and data quality, flagging gaps as they emerge
Use AI-assisted tools (e.g., Claude, Copilot) to speed up triage, root-cause: analysis, scripting, and documentation — while knowing when a problem needs hands-on judgment instead
Write and maintain code in general-purpose languages (Python, Go, JavaScript,: TypeScript, Bash) to automate operational workflows
Partner with global stakeholder teams to drive issues to resolution
What they're looking for
- + years of experience with observability and metrics tooling (e.g.,: Prometheus, Grafana, Datadog, ElasticSearch)
- + years working with data querying and pipelines (e.g., SQL, Airflow, Trino, SQS)
- Working knowledge of network fundamentals and hardware (e.g., Cisco, Palo Alto)
- Experience with Infrastructure as Code