Staff Engineer
Graphcore is a Bristol-based AI chipmaker (IPU accelerators), a SoftBank subsidiary and a Molten Ventures portfolio company.
Maintainer signals as of 8/23/2026
About Graphcore
Graphcore (graphcore.ai) is a British semiconductor company building Intelligence Processing Units (IPUs) for AI workloads. Founded in Bristol in 2016, it was acquired by SoftBank in 2024. It is a portfolio company of Molten Ventures.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will own software engineering efforts across the full software development lifecycle, including implementation, automated testing, integration and production readiness. You will own critical infrastructure, configure and test AI hardware with continuous deployment and infrastructure as code, operate AI systems with datacenter operations engineers and drive corrective actions for system issues.
Requirements
- Have a bachelor's degree or equivalent practical experience in a relevant subject
- Have experience with RESTful API development
- Build deploy and operate containerized workloads using Kubernetes and Docker or Podman
- Have experience managing production Kubernetes clusters and workloads
- Have programming experience with Go
- Have hands-on experience deploying and operating infrastructure using infrastructure as code source code version control and CI/CD automation tools
- Have experience with Redfish for datacenter hardware management telemetry provisioning and control
- Have experience specifying scoping estimating and detailing work plans in an Agile and Scrum framework
- Have strong Linux systems engineering experience including administration automation and Bash and Python scripting
- Have experience with AI coding assistants such as Codex or Claude
- Have experience with Kubernetes operator development and custom resources
- Have experience with high-performance computing environments using SLURM or similar batch workload solutions
- Have experience with virtualized deployments and Open vSwitch KVM or QEMU
- Have experience with distributed object block and file storage such as Ceph
- Have experience with end-to-end deployment automation and CI of containerized services
- Have experience with monitoring and observability solutions such as Grafana Prometheus OpenSearch ElasticSearch Loki Mimir OpenTelemetry Fluentd or Kafka
- Have experience with managed switch configuration such as EOS SONiC or DNOS
- Have experience with PyTorch for AI workloads
- Understand cloud and infrastructure technologies including APIs virtualization networking block storage resource management and monitoring systems
Responsibilities
- Own software engineering efforts across the full software development lifecycle
- Implement automated testing integration and production readiness for the rack management solution
- Drive critical infrastructure issues to resolution
- Configure and test AI hardware and systems using continuous deployment and infrastructure as code
- Maintain and operate AI systems with datacenter operations engineers
- Drive corrective actions for systems that are not operating correctly
Benefits
- Flexible working
- Medical coverage
- Dental coverage
- Vision coverage
- Flexible Spending Accounts
- Health Savings Accounts
- Disability insurance
- Life insurance
- 401(k) retirement plan
- Commuter benefits
- Wellness services
- Employee Assistance Programme
