Safeguards Enforcement Analyst Cyber Harm
AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.
Maintainer signals as of 9/23/2026
Funding history
About Anthropic
Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will review flagged content and accounts, make documented enforcement decisions, and mitigate AI-enabled cyberattacks, malware, exploitation tooling, and related misuse. You will triage severe cases, improve detection quality with technical partners, maintain review standards, and track the evolving cyber threat landscape.
Requirements
- Cybersecurity
- Offensive security
- Exploit development
- Malware analysis
- Vulnerability research
- Content review
- Abuse investigation
- Policy enforcement
- SQL
- Python
- Data analysis
- Threat detection
- Generative AI
- Prompt engineering
Responsibilities
- Review flagged content and accounts and make documented enforcement decisions
- Detect and mitigate AI-enabled cyberattacks, malware creation, exploitation tooling, and harmful cyber operations
- Triage and escalate novel, ambiguous, and high-severity cases
- Provide policy feedback based on enforcement scenarios
- Surface detection-model errors and quality signals to Engineering and Data Science
- Maintain accuracy and consistency across review queues
- Monitor enforcement practices, threat actor tactics, and the cyber threat landscape
Benefits
- Visa sponsorship
- Equity donation matching
- Generous vacation
- Parental leave
- Flexible working hours
