Staff / Senior Machine Learning Engineer Reinforcement Learning

Wayve is a London-headquartered embodied-AI company developing and licensing mapless, vehicle-agnostic driving software for assisted, automated, and robotaxi applications.

London, United Kingdom
About Wayve

Wayve Technologies Ltd. develops the Wayve AI Driver, an end-to-end, data-trained software platform that runs on onboard vehicle compute and native sensors. It is designed for OEM integration across L1 driver assistance through L4 automated driving, without HD maps.

View jobs by Wayve

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will advance reinforcement learning methods for end-to-end driving models. You will design and run large-scale experiments, evaluate policy improvements across simulation and on-road testing, improve reward signals, and productionize successful methods in the shared ML stack.

Requirements

  • Reinforcement learning or sequential decision-making experience
  • Understanding of policy learning, value learning, off-policy learning, function approximation, distribution shift, and learned-objective failure modes
  • Behaviour cloning or reinforcement learning experience
  • Python
  • PyTorch
  • Machine learning training and evaluation systems
  • Experimental design
  • Technical leadership
  • Written communication
  • Verbal communication

Responsibilities

  • Shape and execute the reinforcement learning roadmap
  • Develop and evaluate post-behavior-cloning optimization methods
  • Improve reward models and learning signals
  • Build training and experimentation workflows using driving data
  • Define evidence across offline, simulation, and on-road evaluation
  • Productionize successful methods and mentor colleagues

Benefits

  • Competitive equity package
  • Hybrid working policy
  • Time working from home