The opportunity
Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI.
What you'll do
Lead and grow a 6 engineer team responsible for the reliability, scalability,: and evolution of Figma's observability and AI observability platforms.
Own and evolve the AI observability ecosystem: a privacy-safe telemetry pipeline that aggregates trace, conversation, and model data, runs classification, and lets teams run evals to measure and improve AI feature quality.
Own Figma's core observability stack, including platforms like Datadog,: ensuring high availability, strong data quality, and a healthy signal-to-noise ratio across metrics, logs, and traces.
Set the technical strategy for the instrumentation standards, libraries,: agents, and operators that monitor services across the company.
Explore and ship AI-driven approaches to anomaly detection, root cause: analysis, signal correlation, and operational automation.
Partner across infrastructure, product engineering, finance, and security to: give teams clear visibility into system health at scale, including cost efficiency as the team's scope grows.
What they're looking for
- Coach and develop engineers through career growth, feedback, and technical: leadership, building a culture of ownership and high-quality execution.
- + years of experience leading and growing engineering teams in: infrastructure, observability, platform, or AI systems, with a track record of delivering reliable production systems at scale.
- A strong foundation in distributed systems and a platform mindset: you build systems other teams adopt and rely on, not just for your own team.
- Hands-on depth in observability, AI systems telemetry, or both, whether that: means metrics, logs, and distributed tracing, or the traces, classification, and evals that power AI products.