The opportunity
Crusoe is on a mission to accelerate the abundance of energy and intelligence . As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads.
What you'll do
Run the programs that improve the KPI infrastructure, covering SLO: dashboards, incident follow-up completion rates, on-call burden, deployment velocity, and initiative trackers. Surface where the org needs to improve and drive execution against it. Data should tell the story before anyone has to ask.
Support the Chief of Staff on strategic projects and cross-cutting: initiatives that don't have a natural single owner.
Partner with engineering, SRE, customer success, and data engineering to keep: operational data accurate and consistently reported.
Work toward a single pane of glass that makes the health of the organization: and its deployments easy to understand at a glance.
Own the checks that ensure a quality hand-off between teams, with the: authority to pull in resources needed to deliver compute capacity to customers.
Be the single point of authority to pull in the right teams the moment an: issue surfaces during a deployment, replacing today's back-and-forth, ticket-driven handback.
What they're looking for
- Own the deployment process end to end and evolve it as headcount and site count grow.
- + years in software engineering, technical program management, or a technical: product role, close enough to production systems to know what an SLO breach actually means.
- + years of experience managing teams, with a track record of hiring, developing, and retaining talent
- You've operated in high-growth infrastructure environments where processes: are still being built; ambiguity doesn't paralyze you.