Research Engineer Machine Learning Reinforcement Learning

3 weeks agoSalary: 260K - 630KLondon, UKHybridAiJobs by Anthropic

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

Series F+Recently funded0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/23/2026

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will advance the capabilities and safety of large language models through reinforcement learning research and engineering. You will implement novel approaches, develop agentic models for tool use and open-ended tasks, build training environments and evaluations, and optimize distributed training and evaluation infrastructure.

Requirements

  • Proficiency in Python
  • Experience with async or concurrent programming
  • Experience with machine learning frameworks
  • Industry experience in machine learning research
  • Systems design skills
  • Communication skills

Responsibilities

  • Architect and optimize reinforcement learning infrastructure
  • Design, implement, and test training environments, evaluations, and methodologies for reinforcement learning agents
  • Improve performance through profiling, optimization, benchmarking, caching, and distributed-systems debugging
  • Develop automated testing frameworks and clean APIs
  • Build scalable infrastructure for AI research

Benefits

  • Visa sponsorship
  • Optional equity donation matching
  • Generous vacation
  • Parental leave
  • Flexible working hours
Research Engineer Machine Learning Reinforcement Learning at Anthropic | JobStash