Senior Runtime Engineer
Cerebras builds wafer-scale AI computing systems and a cloud inference platform for training, fine-tuning, and serving AI models.
About Cerebras Systems, Inc.
Cerebras Systems is an AI-infrastructure company founded in 2015. It sells rack-scale wafer-scale computing systems and provides cloud-based, API-accessible AI inference alongside on-premises deployments.
Skills
About the Role
You will design and implement distributed runtime components for large-scale execution workloads. You will optimize data and communication pipelines across compute, memory, storage, and network resources, enable scalable multi-node execution, and resolve performance bottlenecks. You will collaborate on model and hardware optimizations, use profiling tools to diagnose issues, and contribute to system design and roadmap planning.
Requirements
- 3+ years of experience developing high-performance or distributed system software
- Strong C/C++ programming skills
- Expertise in multithreading, memory management, and performance optimization
- Experience with distributed systems, networking, or inter-process communication
- Understanding of data structures, concurrency, and system-level CPU, I/O, and memory management
- Ability to debug, profile, and optimize code from threads to clusters
- Bachelor's, master's, or equivalent experience in computer science, electrical engineering, or a related field
Responsibilities
- Design and implement distributed runtime components for large-scale execution workloads
- Develop and optimize high-performance data and communication pipelines
- Enable scalable execution across multiple compute nodes with high concurrency and minimal bottlenecks
- Collaborate on model architectures, training regimes, and hardware-specific optimizations
- Diagnose and resolve performance issues using profiling and instrumentation tools
- Contribute to system design, architecture reviews, and roadmap planning
