Director or Senior Manager AI Inference Model Scaling
Cerebras builds wafer-scale AI computing systems and a cloud inference platform for training, fine-tuning, and serving AI models.
About Cerebras Systems, Inc.
Cerebras Systems is an AI-infrastructure company founded in 2015. It sells rack-scale wafer-scale computing systems and provides cloud-based, API-accessible AI inference alongside on-premises deployments.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will define the technical vision, organizational strategy, and execution roadmap for a distributed engineering organization. You will lead model compilation, optimization, and high-performance kernel development; hire and mentor engineers; set standards; partner across functions; and deliver model enablement for customer commitments.
Requirements
- BS, MS, or PhD in Computer Science, Computer Engineering, or a related field
- 12+ years building compiler, machine learning systems, or infrastructure software
- 5+ years leading engineering teams
- Compiler infrastructure including LLVM, MLIR, XLA, TVM, or Torch FX
- Graph compilation and optimization
- Python
- C++
- Production software delivery
Responsibilities
- Define the technical roadmap and strategy
- Establish technical direction and engineering standards
- Drive support for emerging LLM architectures and inference workloads
- Hire, mentor, and grow engineering teams
- Develop technical leaders and managers
- Drive organizational planning and investment priorities
- Partner with cloud, machine learning, hardware, product, and customer teams
- Own planning, prioritization, and execution across initiatives
- Drive delivery for strategic customer commitments
