Product Manager Safeguards Account Integrity and Abuse

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

Series F+Recently funded0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/23/2026

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will lead the ideation, design, development, and deployment of safeguards systems and related product experiences. You will define safety evaluations, detections, interventions, and metrics; prioritize technical and business tradeoffs; and align research, engineering, policy, enforcement, and product stakeholders.

Requirements

  • Experience making technical tradeoff decisions
  • Experience collaborating with policy experts, AI or machine-learning researchers, and software engineers
  • Experience building product and engineering strategy across cross-functional teams
  • Experience designing metrics for risk, system performance, and user impact
  • Experience prioritizing rapidly changing product specifications
  • Experience planning, building, launching, and measuring zero-to-one products or systems
  • Ability to communicate complex technical concepts to non-technical audiences
  • 5+ years of product management experience
  • Experience with data, detection, intervention, infrastructure, tools, or evaluations
  • Bachelor’s degree or equivalent education, training, or experience

Responsibilities

  • Design safety-by-design and downstream defenses for AI models and products
  • Write safety evaluations and communicate about safety externally
  • Define problems, options, tradeoffs, and requirements for product delivery
  • Align policy, enforcement, research, engineering, and cross-functional stakeholders
  • Plan mitigations for deployment risks and adversarial misuse
  • Develop metrics for risk, performance, and blind spots

Benefits

  • Equity donation matching
  • Generous vacation
  • Parental leave
  • Flexible working hours
Product Manager Safeguards Account Integrity and Abuse at Anthropic | JobStash