Staff Engineer
Graphcore is a Bristol-based AI chipmaker (IPU accelerators), a SoftBank subsidiary and a Molten Ventures portfolio company.
Maintainer signals as of 8/14/2026
About Graphcore
Graphcore (graphcore.ai) is a British semiconductor company building Intelligence Processing Units (IPUs) for AI workloads. Founded in Bristol in 2016, it was acquired by SoftBank in 2024. It is a portfolio company of Molten Ventures.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will develop and operate the critical interfaces used to manage AI system state and rack hardware. You will own software delivery across the full lifecycle, deploy and test hardware with infrastructure-as-code, operate Kubernetes workloads and AI systems, resolve infrastructure issues, and collaborate with data center operations and engineering teams.
Requirements
- Bachelor's degree or equivalent practical experience in a relevant subject
- Experience with RESTful API development
- Experience building, deploying, and operating containerized workloads using Kubernetes and Docker or Podman
- Experience managing production Kubernetes clusters and workloads
- Programming experience with Go
- Experience deploying and operating infrastructure using infrastructure-as-code, version control, and CI/CD tools
- Experience with Redfish for data center hardware management, telemetry, provisioning, and control
- Experience specifying, scoping, estimating, and detailing work plans in Agile and Scrum frameworks
- Strong Linux systems engineering experience with administration, automation, and Bash and Python scripting
- Experience with Kubernetes operator development is desirable
- Experience with HPC environments using SLURM or similar solutions is desirable
- Experience with virtualized deployments, distributed storage, monitoring, observability, managed switches, or PyTorch is desirable
Responsibilities
- Own software engineering efforts across implementation, automated testing, integration, and production readiness
- Drive critical infrastructure issues to resolution
- Configure and test AI hardware and systems using continuous deployment and infrastructure-as-code
- Work with data center operations engineers to maintain AI systems at peak performance
- Drive corrective actions for systems that are not operating correctly
Benefits
- Flexible working
- Medical coverage
- Dental coverage
- Vision coverage
- Flexible Spending Accounts
- Health Savings Accounts
- Disability insurance
- Life insurance
- 401(k) retirement plan
- Commuter benefits
- Wellness services
- Employee Assistance Programme
