Staff Research Engineer Multi Agent Scaling
2 days agoLeadSalary: 500K - 850KSan Francisco, CA | New York City, NY | Seattle, WAHybridResearchJobs by Anthropic
AnthropicVisit Anthropic website
AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.
Maintainer signals as of 9/25/2026
San Francisco, United States
Funding history
About Anthropic
Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will design, run, and interpret large-scale agent-team experiments. You will build and scale reliable execution systems, develop evaluations and metrics, investigate bottlenecks, debug failures at scale, and enable other research teams to use the platform.
Requirements
- Significant software engineering, machine learning, or research engineering experience
- Experience owning a substantial system, evaluation, benchmark, agent product, or research project end to end
- Quantitative reasoning about complex systems
- Ability to work from vague questions
- Clear written and verbal communication
- Bachelor’s degree or equivalent education, training, or experience
Responsibilities
- Design, run, and interpret large-scale experiments on agent teams
- Investigate performance and efficiency as team size, compute, and task horizon grow
- Build and scale reliable systems for large agent teams
- Design trustworthy evaluations for long-horizon problems
- Build tooling and metrics for understanding agent-team behavior
- Enable research teams to run experiments on the platform and communicate findings clearly
Benefits
- Visa sponsorship efforts
- Optional equity donation matching
- Generous vacation
- Parental leave
- Flexible working hours
