Research Engineer Cybersecurity RL Reinforcement Learning

AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.

Series F+Recently funded0 current maintainers0 active leadsTeam intelligence

Maintainer signals as of 9/24/2026

San Francisco, United States
About Anthropic

Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.

View jobs by Anthropic

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will design and implement reinforcement-learning environments, conduct experiments and evaluations, and deliver work into production training runs. You will develop approaches for secure coding and vulnerability remediation while collaborating with researchers, engineers, and cybersecurity specialists.

Requirements

  • Experience in cybersecurity research
  • Experience with machine learning
  • Strong software engineering skills
  • Ability to balance research exploration with engineering implementation

Responsibilities

  • Design and implement reinforcement learning environments
  • Conduct experiments and evaluations
  • Deliver work into production training runs
  • Develop secure coding and vulnerability remediation capabilities
  • Collaborate with researchers, engineers, and cybersecurity specialists

Benefits

  • Optional equity donation matching
  • Generous vacation
  • Parental leave
  • Flexible working hours
Research Engineer Cybersecurity RL Reinforcement Learning at Anthropic | JobStash