Software Engineer Cluster Deployment

Cerebras builds wafer-scale AI computing systems and a cloud inference platform for training, fine-tuning, and serving AI models.

Sunnyvale, California, United States
About Cerebras Systems, Inc.

Cerebras Systems is an AI-infrastructure company founded in 2015. It sells rack-scale wafer-scale computing systems and provides cloud-based, API-accessible AI inference alongside on-premises deployments.

View jobs by Cerebras Systems, Inc.

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will build and maintain automation for cluster provisioning, configuration, validation, and handoff. You will turn manual steps into tested workflows, participate in deployments, troubleshoot infrastructure, and add health checks and observability. You will contribute to infrastructure-as-code and GitOps workflows with partner teams.

Requirements

  • 2+ years of mid- to large-scale data-center deployment experience
  • Python and Bash
  • Linux command-line, processes, filesystems, and disk troubleshooting
  • Git, branching, commits, pull requests, and code review
  • Technical degree or equivalent practical experience

Responsibilities

  • Develop and maintain deployment automation
  • Turn manual deployment steps into tested pushbutton workflows
  • Participate in hands-on cluster deployments
  • Troubleshoot Linux, servers, networking, storage, Kubernetes, and connectivity
  • Contribute to infrastructure-as-code and GitOps workflows
  • Add health checks, observability, dashboards, and validation logic
  • Partner with infrastructure, networking, security, and operations teams
Software Engineer Cluster Deployment at Cerebras Systems, Inc. | JobStash