Staff+ Site Reliability Engineer Safeguards ML Infra
AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.
Maintainer signals as of 9/23/2026
Funding history
About Anthropic
Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will configure, verify, and deploy safety safeguards for model launches across production platforms. You will operate and improve classifier rollouts, canary changes, validate deployments, identify configuration drift, and retain rollback authority when needed. You will turn runbooks and manual checks into automated tooling and pipelines, maintain deployment provenance, and participate in on-call incident response.
Requirements
- Production change-management experience at scale
- Experience with deployment pipelines, configuration management, and canary analysis
- Experience leading high-stakes releases or incident response
- On-call experience for production systems
- Experience deploying and operating cloud platforms at scale
- Python proficiency
Responsibilities
- Configure and verify safeguards for model releases
- Deploy new safety classifiers through canary rollouts and post-deployment validation
- Verify safeguards across deployment platforms and eliminate configuration drift
- Automate launch runbooks, validation checks, and deployment workflows
- Build and maintain a safeguards registry with deployment provenance
- Participate in on-call rotations, incident response, and model provisioning
Benefits
- Optional equity donation matching
- Generous vacation
- Parental leave
- Flexible working hours
