The opportunity
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
What you'll do
Own end-to-end technical execution for server systems and network equipment: in Cerebras clusters, including NPIs, platform refreshes, and major component or configuration changes.
Drive requirements gathering and technical trade-off decisions, converting: inputs into executable plans with clear milestones, readiness gates, and cross-functional deliverables.
Represent Cluster Architecture in executive reviews, OKR cycles, and leadership or customer forums as needed.
Build and manage integrated execution plans across vendors and internal: teams, tracking dependencies, critical paths, and risks.
Lead OEM/ODM, switch-vendor, and component-vendor engagements, including: RFI/RFP activities, technical evaluations, samples, escalations, and roadmap alignment.
Partner with Compute, Server Platform, and Network Architects to translate: architectural direction into platform decisions, qualification plans, acceptance criteria, and rollout strategies.
What they're looking for
- Lead NPI execution, qualification, and release readiness, including: lab/staging validation, regression tracking, issue resolution, and go/no-go decisions.
- Step into execution gaps as needed to drive technical issues to closure and: keep server and network programs moving.
- Own risk and change management into production, including versioning, rollout: sequencing, and stakeholder communication.
- Ensure operational readiness with deployment and fleet teams and maintain: alignment with rack and physical datacenter owners on power, cooling, space, and cabling constraints.