Member of Technical Staff Safety

Reflection is an AI research lab building open frontier models and a full AI stack for developers, enterprises, and public-sector users.

New York, United States
About Reflection

Reflection develops open-weight AI models, open-source software for customizing and running agents, AI-factory infrastructure, and related solutions. Its current research emphasizes large language models, reinforcement learning, and agentic reasoning.

View jobs by Reflection

Skills

About the Role

You will own red-teaming and adversarial evaluation pipelines for AI models. You will identify security, misuse, and alignment failures; develop automated safety benchmarks; implement jailbreaking techniques and defenses; translate findings into guardrails; and validate releases against risk thresholds.

Requirements

  • Graduate degree in Computer Science Machine Learning or a related discipline or equivalent AI safety experience
  • Knowledge of LLM safety adversarial attacks red-teaming methodologies and interpretability
  • Software engineering experience building automated evaluation pipelines or large-scale ML systems
  • Ability to make high-stakes model release and safety-threshold decisions

Responsibilities

  • Own red-teaming and adversarial evaluation pipelines
  • Probe models for security misuse and alignment failure modes
  • Translate safety findings into concrete guardrails
  • Validate releases against safety risk thresholds
  • Develop scalable automated safety benchmarks
  • Research and implement jailbreaking techniques and defenses

Benefits

  • Stock options
  • Medical dental vision and life insurance
  • Annual wellness allowance
  • Daily office lunch and dinner
  • 22 weeks of paid parental leave
  • Unlimited paid time off in the United States
  • 30 days of vacation in the United Kingdom
  • Visa sponsorship support
  • Regular off-sites happy hours and team celebrations
Member of Technical Staff Safety at Reflection | JobStash