The opportunity
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
What you'll do
Provision and configure network devices, firewalls, and servers in data: centers by following documented processes and using purpose-built provisioning frameworks and automation tools.
Troubleshoot and resolve server, network, configuration, automation, and: connectivity issues encountered during infrastructure provisioning and service bring-up.
Collaborate with engineering teams, data center technicians, cabling vendors,: and rack integrators to coordinate infrastructure deployment and ensure timely site readiness.
Support network, power, and mechanical commissioning, including testing and: validation of network connectivity, device configurations, server readiness, and infrastructure dependencies.
Develop and continuously improve provisioning processes, tooling, and: automation to enable reliable, repeatable, and scalable deployment across multiple sites in parallel.
Perform post-deployment validation and support integration with monitoring,: alerting, and operational tooling to ensure infrastructure meets production-readiness requirements.
What they're looking for
- Coordinate the successful handoff of deployed infrastructure to Operations &: Maintenance (O&M) teams - Cluster Operations, Network Operations, Site Operations.
- Document deployment procedures, troubleshooting findings, lessons learned,: and process improvements to increase deployment efficiency and reliability.