Staff+ Software Engineer Distributed Systems

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

Series F+Recently funded0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/23/2026

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will lead complex infrastructure projects and build systems that detect unwanted model behavior, support enforcement, and surface analyst dashboards. You will deploy monitoring across clouds, harden sandboxed agent runtimes, manage platform capacity and service objectives, and partner on trusted safety, privacy, and legal controls.

Requirements

  • Have a bachelor's degree in computer science, software engineering, or comparable experience
  • Be proficient in Python, distributed systems, and infrastructure as code
  • Work across multiple cloud providers or build provider-agnostic infrastructure
  • Build and operate large-scale distributed infrastructure
  • Have experience with sandboxing, isolation technologies, containers, microVMs, network policy, or Rust
  • Explain complex technical concepts to non-technical stakeholders
  • Build systems for sensitive or regulated data
  • Have experience with LLM-based agents in production
  • Have experience with integrity, spam, fraud, or abuse detection and mitigation

Responsibilities

  • Lead complex multi-month infrastructure projects
  • Develop monitoring systems for unwanted model behavior and automated enforcement
  • Surface monitoring data in internal analyst dashboards
  • Deploy monitoring systems across multiple clouds
  • Maintain deployment pipelines, smoke tests, observability, and alerting
  • Design and harden sandboxed agent runtimes
  • Manage cost and capacity for long-running agent workloads
  • Set and meet platform service-level objectives
  • Partner with engineering, research, security, privacy, and legal stakeholders

Benefits

  • Equity donation matching
  • Vacation leave
  • Parental leave
  • Flexible working hours