Anthropic Fellows Program ML Systems and Reinforcement Learning

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

About the Role

You will complete a four-month empirical research project with mentorship. You may build accelerator infrastructure, synthetic-data or environment pipelines, model-based tools, and reinforcement-learning environments. You will research training data, generalization, and RL algorithms, and work toward a public output.

Requirements

  • Fluency in Python programming
  • Availability to work full-time for the Fellows Program
  • Strong software engineering skills
  • Experience building complex ML systems
  • Work authorization in the US, UK, or Canada
  • Location in the US, UK, or Canada during the program

Responsibilities

  • Conduct an empirical research project
  • Build ML systems and infrastructure
  • Build synthetic-data or environment pipelines
  • Build model-based tools for training-data analysis
  • Create reinforcement-learning environments
  • Research and implement reinforcement-learning algorithms
  • Produce a public research output

Benefits

  • Direct mentorship from Anthropic researchers
  • Access to a shared workspace
  • Connection to the AI safety and security research community
  • Funding for compute and other research expenses

Hiring Process

Initial application and reference check → technical assessments and interviews → research discussion

Anthropic Fellows Program ML Systems and Reinforcement Learning at Anthropic | JobStash