The opportunity
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers.
What you'll do
Define and own Lambda’s data center engineering standards for high-density AI: and GPU infrastructure environments.
Create repeatable reference architectures, design patterns, technical: specifications, and deployment standards that improve speed, quality, reliability, and consistency across data center builds.
Develop the engineering “operating system” for data center scale, including: standards libraries, design review processes, decision records, quality gates, commissioning criteria, acceptance checklists, exception processes, and operational handoff frameworks.
Translate lessons learned from individual deployments into reusable: mechanisms that make future deployments safer, faster, and more predictable.
Establish technical standards across areas such as power distribution,: cooling, liquid cooling readiness, rack integration, structured cabling, fiber management, network rooms, out-of-band management, telemetry, DCIM/BMS integrations, physical security, maintainability, and operational readiness.
Lead cross-functional architecture reviews for new data center designs,: expansions, retrofits, and infrastructure programs.
What they're looking for
- Identify organizational bottlenecks and design systems, tools, and workflows: that reduce ambiguity, improve accountability, and help teams execute at scale.
- Build governance mechanisms for standards adoption, including exception: management, risk reviews, lifecycle ownership, metrics, and continuous improvement loops.
- Act as a technical advisor to senior engineering and infrastructure: leadership on data center strategy, design tradeoffs, resiliency, operational risk, and scalability.
- Mentor engineers and technical leaders across the organization by raising the: bar for systems thinking, documentation, design rigor, and operational excellence.