Senior Software Engineer, StoragePosted today$196K

The opportunity

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences…

What you'll do

  • Database Kernel: The Kernel team owns performance, reliability, and stability for Airbnb's NewSQL engine at scale, spanning storage, query execution, consensus, CDC, and central routing and shard management. In this role you'll do deep technical investigation of database internals and SQL latency and availability issues, and partner closely with our open-source database vendor on root cause analysis and upstream contributions.

  • Database Adoption: The Adoption team builds the infrastructure that moves Airbnb's data workloads onto NewSQL: eventually consistent two-way replication between databases, fully automated failover and failback, and a migration operator. In this role you'll work across Kernel, SRE, and the ORM, KV, and SQL interface teams to make onboarding and migration smooth for internal customers.

  • Dig into the query and storage layers to explain latency differences between: B-tree and LSM-based storage engines.

  • Design and implement traffic prioritization at the storage layer for connection pooling.

  • Validate and safely roll out a new CDC architecture while guaranteeing full: transactional correctness across shards.

  • Build the frameworks and tooling around our NewSQL database: monitoring, observability, permissions, and service discovery integration.

What they're looking for

  • + years building and operating large-scale distributed systems: storage, data ingestion, backup and restore, or streaming.
  • Hands-on experience with database or storage internals: query execution, storage engines, replication, or consensus.
  • Ability to own and dive deeply into a complex codebase, ideally an open-source one.
  • Experience maintaining, analyzing, and debugging high-severity production incidents.