Software Engineer II, Managed Platform ServicesActive$140K–$165K

The opportunity

Crusoe is on a mission to accelerate the abundance of energy and intelligence . As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads.

What you'll do

  • Building Foundational Infrastructure: Build and scale core infrastructure services that manage critical resources within our cloud platform. This involves designing, developing, and deploying robust and reliable systems from the ground up.

  • Scalable Design: Design highly scalable, durable, and reliable platform services that prioritize ease of use.

  • Cross Functional Collaboration: Lead projects that require collaborating with engineering, cloud support, site reliability, and product teams to assess tools, frameworks, and solutions that align with both customer and operational needs.

  • Innovation: Implement features that differentiate Crusoe Cloud, focusing on: operational efficiency, low-touch adoption, turn-key AI services, and scalability.

  • Component ownership & refactoring: Own the end to end development and maintenance of specific software modules and service components. Proactively identify and execute refactoring opportunities to improve code modularity, readability, and long term maintainability.

  • Full lifecycle feature development: Design, develop, and deploy complex features that span the entire software development lifecycle. Translate high level business requirements into technical specifications, ensuring solutions are scalable and integrated seamlessly into the existing cloud architecture.

What they're looking for

  • Architectural design for new features: Lead the design phase for new functionality by creating simple, elegant solutions for difficult technical problems. Utilize advanced knowledge of software design patterns to minimize system complexity and prevent technical debt.
  • Operational excellence & risk mitigation: Identify and mitigate visible technical risks and roadblocks within project workstreams. Enhance the team’s operational health by automating manual processes and improving monitoring, logging, and alerting systems for production services.
  • Advanced troubleshooting & root cause analysis: Independently diagnose and resolve intricate software failures and performance bottlenecks. Conduct root cause analysis on production incidents and implement systemic fixes to prevent recurrence across the component's lifecycle.
  • Mentorship & knowledge sharing: Accelerate the onboarding of new teammates by providing technical training on team-specific software and development processes. Act as a peer reviewer for code review, offering constructive feedback to ensure high quality output across the team.