Data Center Provisioning Engineer
Cerebras builds wafer-scale AI computing systems and a cloud inference platform for training, fine-tuning, and serving AI models.
About Cerebras Systems, Inc.
Cerebras Systems is an AI-infrastructure company founded in 2015. It sells rack-scale wafer-scale computing systems and provides cloud-based, API-accessible AI inference alongside on-premises deployments.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will provision, configure, commission, and validate data center network infrastructure, servers, and related systems. You will troubleshoot provisioning and connectivity issues, coordinate site readiness with technical partners, improve deployment automation, validate production readiness, and document procedures and lessons learned.
Requirements
- Bachelor's degree in Electrical Engineering, Computer Science, Computer Engineering, or a related technical field
- At least 5 years of relevant infrastructure provisioning, data center, or SRE experience
- Server and network device provisioning and troubleshooting experience
- Python and shell scripting proficiency
- Experience with AWS, GCP, or Azure
- Kubernetes deployment, scaling, troubleshooting, and cluster management experience
- Terraform and Ansible experience
- Knowledge of PXE, DNS, DHCP, TCP/IP, load balancing, VPCs, firewalls, VLANs, LACP, BGP, and OSPF
- Experience upgrading server, switch, and PDU firmware
Responsibilities
- Provision and configure network devices, firewalls, and servers
- Troubleshoot server, network, configuration, automation, and connectivity issues
- Coordinate infrastructure deployment and site readiness with engineering, technicians, vendors, and integrators
- Support network, power, and mechanical commissioning
- Improve provisioning processes, tooling, and automation
- Validate deployed infrastructure and integrate it with monitoring and operational tooling
- Hand off deployed infrastructure to operations and maintenance teams
- Document deployment procedures, troubleshooting findings, and process improvements
