Data Center Engineer Reliability and Infrastructure Management Compute Supply

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

Series F+Recently funded0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/25/2026

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will model and manage data center power, cooling, availability, topology, and load behavior. You will work with developers, operators, cloud providers, and vendors to improve designs, establish performance requirements, forecast capacity, and support commissioning and incidents.

Requirements

  • Bachelor's degree in Electrical Engineering, Mechanical Engineering, Power Systems, Reliability Engineering, Controls Engineering, or a related field
  • 5+ years of experience in data center infrastructure, facility engineering, or reliability engineering
  • Experience with data center power distribution, cooling architectures, and failure-mode management
  • Experience building reliability or availability models, power or capacity models, or software-based power-management or control systems
  • Knowledge of SCADA, BMS, EPMS, telemetry pipelines, and control systems
  • Cross-functional collaboration experience across hardware, software, and facilities teams

Responsibilities

  • Own power and cooling topology data and validate telemetry
  • Define load-management behavior under failures
  • Build and audit availability models for electrical and mechanical systems
  • Define cloud availability zones and failure domains
  • Build power, energy, capacity, and load-forecasting models
  • Conduct technical diligence with developers, operators, cloud providers, and chip vendors
  • Drive design improvements and SLA-grade performance

Benefits

  • Visa sponsorship
  • Optional equity donation matching
  • Generous vacation
  • Parental leave
  • Flexible working hours