Threat Intelligence Manager Model Exploitation and Fraud

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

Series F+Recently funded0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/24/2026

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will set the strategy and priorities for model exploitation and fraud investigations. You will hire and develop technical investigators, direct complex casework, improve high-volume triage, build fraud and scam playbooks, and turn findings into enforcement and product mitigations. You will also lead intelligence sharing with government and industry partners and brief leadership on threats.

Requirements

  • Experience leading investigative, fraud, platform integrity, or threat intelligence teams
  • Domain knowledge of scaled abuse, fraud, account abuse, unauthorized access, or platform exploitation economics
  • SQL and Python proficiency
  • Experience overseeing investigations across surface, deep, and dark web environments
  • Working familiarity with large language models and model exploitation at scale
  • Experience building processes, detection systems, or programs from scratch
  • Executive, engineering, and external-partner communication skills

Responsibilities

  • Own strategy, priorities, and outcomes for the Model Exploitation and Fraud mission area
  • Hire, manage, and develop technical threat investigators
  • Direct complex investigations into model distillation, unauthorized access, account abuse, and fraud networks
  • Redesign triage and build abuse signals, clustering, and agentic investigation workflows
  • Build fraud and scam detection and investigation playbooks
  • Lead intelligence sharing with U.S. government partners and industry peers
  • Convert findings into bans, product mitigations, and safety-by-design improvements
  • Define team metrics and brief leadership on the threat landscape

Benefits

  • Optional equity donation matching
  • Generous vacation
  • Parental leave
  • Flexible working hours
Threat Intelligence Manager Model Exploitation and Fraud at Anthropic | JobStash