Misuse Red Team Research Engineer Research Scientist

UK government research organisation that evaluates advanced AI risks and develops and tests mitigations to inform governments.

Distributed
About AI Security Institute

The AI Security Institute (AISI) is a research organisation within the UK Department for Science, Innovation and Technology. It conducts technical research, evaluates leading AI systems, develops risk mitigations, and shares evaluation infrastructure such as Inspect.

View jobs by AI Security Institute

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will develop, run, and evaluate automated attacks and safeguards for frontier AI systems. You will build monitoring benchmarks, investigate data-poisoning attacks and defences, conduct adversarial testing, and produce actionable reports for safeguard developers.

Requirements

  • Large language model research
  • Large language model training
  • Fine-tuning
  • Model evaluation
  • AI safety research
  • Machine learning
  • PyTorch
  • Inspect
  • Research code
  • Peer-reviewed publication

Responsibilities

  • Design, build, run, and evaluate automated attacks and safeguard evaluations
  • Build benchmarks for asynchronous monitoring of misuse and jailbreak development
  • Investigate attacks and defences for LLM data poisoning and backdoors
  • Conduct adversarial testing of frontier AI safeguards
  • Produce actionable reports for safeguard developers

Benefits

  • Pre-release access to frontier models and compute
  • Learning and development stipend
  • Conference and external collaboration funding
  • Hybrid working
  • Occasional remote work abroad
  • Work-from-home equipment stipend
  • At least 25 days of annual leave
  • 8 public holidays
  • Additional team-wide breaks
  • 3 volunteering days
  • Paid parental leave
  • Employer pension contribution
  • Cycling, donation, retail, and gym discounts

Hiring Process

Initial assessment; initial screening call; research interview; technical assessment; behavioural interview; final interview with members of the senior leadership team.