The opportunity
Build ingestion pipelines to consume 1M+ RPS streams of logs, metrics, and other telemetry
What you'll do
Build ingestion pipelines to consume 1M+ RPS streams of logs, metrics, and other telemetry
Build scalable, fault tolerant alerting engines for notifying users, in real-time, of threshold breaches
Craft rich backend observability APIs, working with product to build amazing: experiences for instantly grokking their application
Provide APIs to access realtime log/metrics streams to be consumed by the Dashboard and Product Teams
Build Golang/Rust GRPC services from scratch capable of supporting tens of: thousands of users, and the million+ to come.
Define infrastructure that can be torn down, failed over, and reconstituted: from scratch using principle of immutable infrastructure using Terraform and Ansible.
What they're looking for
- Write Engineering Requirement Documents to take something from idea, to: defined tasks, to implementation, to monitoring it’s success.
- Interface with our TypeScript and GraphQL edge to expose your microservice: APIs for both internal and potentially external consumption
- A strong understanding of distributed systems. You enjoy building fault: tolerant, resilient, and scalable services
- Interests in VictoriaMetrics, ClickHouse, and other systems for building: observability stacks from the ground up