Distributed Systems Engineer - Data Platform (Delivery, Database, Retrieval)Active$150K–$206K
The opportunity
Cloudflare’s Data Org builds the systems that ingest, process, store, and retrieve data generated across our global network. These systems power the logs, analytics, and alerts that give customers near real-time visibility into the security, reliability, and performance of their Internet properties.
What you'll do
Design, build, and operate reliable distributed systems that process very large volumes of data.
Develop and optimise backend services and data platform components, primarily in Go.
Improve data integrity, delivery latency, query performance, availability, and operational efficiency.
Build systems that handle failure safely through techniques such as: backpressure, retries, idempotency, replication, and graceful degradation.
Define service-level objectives, monitor production health, plan capacity, and respond to incidents.
Diagnose performance bottlenecks across ingestion, processing, storage, and retrieval.
What they're looking for
- At least three years of professional software engineering experience building: backend, infrastructure, database, or distributed-data systems.
- Strong Go programming skills, or substantial systems programming experience: (Rust, C++) with the ability to become productive in Go quickly.
- Experience building and operating production services, including deployment,: monitoring, capacity planning, incident response, and debugging.
- A solid understanding of distributed-systems concepts, including concurrency,: consistency, fault tolerance, partitioning, and data integrity.
- Experience with high-throughput or low-latency data processing.
- Hands-on experience with observability tools such as Prometheus and Grafana,: including monitoring high-cardinality workloads.
- Strong analytical and troubleshooting skills, particularly for complex production problems.
- Clear communication and an ability to collaborate across engineering and product teams.