Staff Software Engineer, Deployment PlatformActive

The opportunity

The Core Change Management group is responsible for the systems that let every Stripe engineer ship code, configuration, and infrastructure changes safely and at high velocity. You will be embedded primarily on the Service Deployments team — the owners of Stripe's end-to-end…

What you'll do

  • Own end-to-end technical delivery of large, ambiguous infrastructure projects: — from initial design through production launch and long-term reliability. Author the design, sequence the work, unblock the team, and shepherd projects to landed impact.

  • Architect the next generation of Stripe's deployment platform. Lead technical: design of the deployment orchestrator's evolution — including multi-service dependency-aware autodeploy pipelines, Kubernetes-native deployment primitives, and fleetwide container migration — defining the API contracts, rollout strategies, and operational model that hundreds of teams depend on.

  • Extend deploy anomaly detection. Evolve blue-green traffic analysis: extend coverage to earlier traffic-split stages, design API/method-based regression detection, and build a self-service onboarding system that makes anomaly detection the default for all supported service types.

  • Own reliability and operational excellence for the deployment platform. Lead: incident response; systematically reduce operational toil; and make reliability, security, and maintainability first-class properties of the systems you own.

  • Build deployment event infrastructure. Own the deployment notification and: event-publishing architecture — designing the event schema, durability model, and integration contracts that downstream systems rely on for observability and automation.

  • Collaborate across Core Change Management. Partner with Resource Automation: on projects that span deployment orchestration and cloud resource management (IAM, account provisioning, infrastructure automation), with Feature Deployments on change-safety tooling (feature flags, configuration management, change audit logs) that integrates with or depends on the deployment pipeline, and with the service mesh team on routing capabilities that enable advanced deployment patterns such as canary rollouts and merchant-priority traffic shaping.

What they're looking for

  • + years of professional software engineering experience , with a demonstrated: track record of designing and shipping production infrastructure systems of significant scale and complexity.
  • Proven ability to lead large, ambiguous infrastructure projects end-to-end —: from technical design through delivery — including managing cross-team dependencies and coordinating migrations across many consuming teams.
  • Deep expertise in distributed systems and deployment orchestration: strong foundations in how services are built, scheduled, and operated at scale, including rollout strategies, staged delivery, and failure modes.
  • Hands-on experience with Kubernetes and container-based deployments ,: including service lifecycle management, workload scheduling, and the operational challenges of migrating large fleets from VM-based to containerized infrastructure.
  • Strong background in service reliability and operational excellence: demonstrated ability to lead incident response, reduce toil, and build systems that are reliable, debuggable, and maintainable by a team.
  • Track record of broad technical impact across multiple large systems: fluency across a complex codebase, force-multiplier effect through code review and mentorship, and the ability to set technical direction for a team rather than just execute within it.