Anthropic Fellows Program AI Safety and Security

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

About the Role

You will conduct a four-month empirical research project using external infrastructure, such as open-source models and public APIs. You will work toward a public research output, participate in project selection and mentor matching, and contribute to AI safety or AI security research.

Requirements

  • Fluency in Python programming
  • Availability to work full-time for the Fellows Program
  • Work authorization in the US, UK, or Canada
  • Location in the US, UK, or Canada during the program

Responsibilities

  • Conduct an empirical research project aligned with research priorities
  • Use external infrastructure, open-source models, and public APIs
  • Produce a public research output
  • Participate in project selection and mentor matching
  • Conduct AI safety or AI security research

Benefits

  • Direct mentorship from Anthropic researchers
  • Access to a shared workspace
  • Connection to the AI safety and security research community
  • Funding for compute and other research expenses

Hiring Process

Initial application and reference check → technical assessments and interviews → research discussion

Anthropic Fellows Program AI Safety and Security at Anthropic | JobStash