Safeguards Enforcement Analyst User Well-being

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

Series F+Recently funded0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/23/2026

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will design, evaluate, and monitor mental-health interventions and detection systems. You will review flagged content, improve enforcement workflows, translate clinical and policy guidance into measurable criteria, and work with technical and policy partners on crisis-resource features and policy gaps.

Requirements

  • Trust and safety, product policy, content moderation, or related experience involving mental-health or well-being harms
  • Experience designing experiments, evaluations, or measurement studies
  • Experience translating policy into rubrics, review guidelines, or classification criteria
  • Content-review operations, quality assurance, and workflow-management experience
  • SQL or other data-analysis tools
  • Generative AI product and prompt-writing experience
  • Risk analysis and cross-functional communication
  • Knowledge of scaled content-moderation policy implementation
  • Sound judgment in high-consequence cases

Responsibilities

  • Design and execute interventions, define metrics, and curate evaluation datasets
  • Build, tune, and validate detection models with Engineering and Data Science
  • Monitor intervention and detection-system performance
  • Review flagged content to improve enforcement and policy
  • Develop crisis-resource features and referral pathways
  • Provide feedback on policy gaps based on real scenarios
  • Apply emerging AI policy and mental-health research to workflows

Benefits

  • Optional equity donation matching
  • Generous vacation
  • Parental leave
  • Flexible working hours
Safeguards Enforcement Analyst User Well-being at Anthropic | JobStash