# Staff / Senior Machine Learning Engineer, Reinforcement Learning at Wayve

- Company: Wayve
- What the company does: Backed by SoftBank VF and Air Street.
- Company website: https://wayve.ai/
- Type: Startups (AI role)
- Level: Senior
- Location: Sunnyvale
- Work setup: On-site
- Posted: 2026-09-10
- Apply by: 2026-10-25
- Apply: https://wayve.firststage.co/jobs?gh_jid=8795692002
- Page: https://www.1752.vc/careers/jobs/wayve-staff-senior-machine-learning-engineer-reinforcement-learning/

## About the role

As a Senior / Staff Machine Learning Engineer in Wayve's AV Core organisation, you will advance reinforcement learning methods for end-to-end driving models. You will identify where learning from reward or feedback can improve beyond behavior cloning, then take promising ideas from design through large-scale experiments, rigorous evaluation, and integration into our best driving models.

## What they're looking for

- Shape and execute the reinforcement learning roadmap for Driving Core / Core Model Safety, selecting problems and methods against clear behavioral gaps and measurable success criteria
- Help improve the reward models and related learning signals used to train and evaluate driving policies, working with partner teams to strengthen their quality, scalability, and downstream usefulness
- Build robust training and experimentation workflows using large-scale driving data, diagnose distribution shift, objective misspecification, optimization instability, and data or evaluation bias
- Define evidence across offline metrics, open-loop tests, closed-loop simulation, and on-road evaluation, and distinguish genuine policy improvement from benchmark overfitting
- Productionize successful methods in the shared ML stack, communicate decisions and results clearly, and raise the technical bar through design reviews, code reviews, and mentoring

Tags: AV Engineering
