Staff Software Engineer, Account Creation

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

Series F+Recently funded0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/24/2026

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will develop monitoring and abuse-detection systems, enable automated enforcement and analyst review, identify abuse patterns for research teams, and build reliable multilayered defenses that improve safety mechanisms in real time at scale.

Requirements

  • Proficiency in Python and TypeScript
  • Ability to work across the stack
  • Experience with integrity, spam, fraud, or abuse detection and mitigation
  • Experience building trust and safety detection mechanisms for AI/ML systems
  • Experience with prompt engineering, jailbreak attacks, and adversarial inputs
  • Experience building custom internal tooling with operational teams

Responsibilities

  • Develop monitoring systems to detect unwanted behaviors from API partners
  • Enable automated enforcement actions and analyst review through internal dashboards
  • Build abuse-detection mechanisms and infrastructure
  • Surface abuse patterns to research teams
  • Build robust multilayered defenses for real-time safety improvements

Benefits

  • Equity donation matching
  • Generous vacation
  • Parental leave
  • Flexible working hours