Staff / Senior Machine Learning Engineer Reinforcement Learning
Wayve is a London-headquartered embodied-AI company developing and licensing mapless, vehicle-agnostic driving software for assisted, automated, and robotaxi applications.
About Wayve
Wayve Technologies Ltd. develops the Wayve AI Driver, an end-to-end, data-trained software platform that runs on onboard vehicle compute and native sensors. It is designed for OEM integration across L1 driver assistance through L4 automated driving, without HD maps.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will advance reinforcement learning methods for end-to-end driving models. You will design and run large-scale experiments, evaluate policy improvements across simulation and on-road testing, improve reward signals, and productionize successful methods in the shared ML stack.
Requirements
- Reinforcement learning or sequential decision-making experience
- Understanding of policy learning, value learning, off-policy learning, function approximation, distribution shift, and learned-objective failure modes
- Behaviour cloning or reinforcement learning experience
- Python
- PyTorch
- Machine learning training and evaluation systems
- Experimental design
- Technical leadership
- Written communication
- Verbal communication
Responsibilities
- Shape and execute the reinforcement learning roadmap
- Develop and evaluate post-behavior-cloning optimization methods
- Improve reward models and learning signals
- Build training and experimentation workflows using driving data
- Define evidence across offline, simulation, and on-road evaluation
- Productionize successful methods and mentor colleagues
Benefits
- Competitive equity package
- Hybrid working policy
- Time working from home
