Compute Server Platform Architect
Cerebras Systems, Inc.Visit Cerebras Systems, Inc. website
Cerebras builds wafer-scale AI computing systems and a cloud inference platform for training, fine-tuning, and serving AI models.
Sunnyvale, California, United States
About Cerebras Systems, Inc.
Cerebras Systems is an AI-infrastructure company founded in 2015. It sells rack-scale wafer-scale computing systems and provides cloud-based, API-accessible AI inference alongside on-premises deployments.
Skills
About the Role
You will own server-side platform architecture for AI clusters. You will define server configurations, capacity formulas, CPU, memory, PCIe, networking, storage, and firmware baselines. You will model and benchmark performance, lead vendor engagements, establish qualification criteria, and troubleshoot deployment regressions.
Requirements
- PhD and 8+ years of industry experience, or BS or MS and 10+ years of industry experience
- 5+ years of server platform architecture, systems performance engineering, or large-scale infrastructure design
- x86 server architecture
- Linux systems
- NIC behavior
- RDMA
- RoCE
- NVMe
- Capacity and performance modeling
- Benchmarking
- Vendor platform evaluation
- C
- C++
- Python
Responsibilities
- Own architecture for cluster server roles, configurations, and lifecycle strategy
- Define server formulas, capacity planning, and headroom policy
- Specify CPU, memory, PCIe, NIC, and NVMe platform configurations
- Translate runtime flows into hardware requirements
- Develop and validate performance and scaling models
- Define operating system, BIOS, firmware, and driver baselines
- Evaluate emerging server technologies
- Lead vendor engagements and qualification efforts
- Support deployment debugging and root-cause analysis
