Principal Datacenter Technologist

Graphcore is a SoftBank-owned AI-compute company developing AI processors, systems, and software for machine-learning workloads.

Series ERecently funded0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/25/2026

Bristol, United Kingdom
About Graphcore

Graphcore develops Intelligence Processing Units (IPUs) and the Poplar SDK for building and running machine-learning applications. It continues operating under the Graphcore name as a wholly owned SoftBank Group subsidiary.

View jobs by Graphcore

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will define operating models and new-product-introduction workflows for data center technologies. You will establish deployment, service, diagnostics, telemetry, and readiness processes; resolve complex platform issues; guide cross-functional work; and communicate technical risks and recommendations.

Requirements

  • Bachelor's degree or equivalent practical experience in a relevant technical field
  • Extensive data center infrastructure, operations, platform deployment, or related experience
  • Experience defining new-product-introduction processes and readiness criteria
  • Systems knowledge across compute, networking, storage, memory, firmware, operating systems, racks, power, cooling, and data center operations
  • Hands-on Linux scripting or automation experience
  • Experience with failure analysis and validation of complex data center equipment
  • Ability to lead multidisciplinary technical work without direct authority

Responsibilities

  • Own the technical operating model and new-product-introduction framework
  • Define workflows for installation, configuration, validation, deployment, service, repair, upgrades, and sustained operation
  • Translate architecture requirements into data center readiness criteria
  • Create runbooks, interface definitions, ownership models, acceptance criteria, and escalation paths
  • Lead operational readiness reviews and identify gaps in tooling, diagnostics, telemetry, serviceability, and training
  • Serve as a senior escalation point for complex platform issues
  • Guide deployment of automated diagnostic and telemetry systems
  • Establish quality and operational metrics and feed findings into architecture and deployment decisions
  • Coordinate resolution across hardware, firmware, software, networking, facilities, and site operations
  • Mentor engineers and communicate readiness, risks, tradeoffs, and recommendations

Benefits

  • Medical, dental, and vision coverage
  • Mental health, wellness, and employee-assistance resources
  • Retirement savings benefits and company contributions
  • Paid vacation, sick time, holidays, and parental or family leave
  • Life insurance and short-term or long-term disability coverage
  • Flexible working hours and hybrid working arrangements where compatible
  • Office amenities and team-led activities