Wayve
Machine Learning Engineer, Performance Tooling

How your CV stacks up
Upload your CV to see how well it fits this job role
?%
The role
Wayve is building autonomous driving technology that runs on real vehicles. Getting our models onto embedded hardware — correctly, quickly, and reproducibly — is one of the hardest problems between research and product.
As a ML Compiler Engineer, you will own the compilation pipeline that makes that possible. You will build and extend Wayve's ML compiler end-to-end: designing passes, integrating with vendor toolchains like NVIDIA TensorRT and Qualcomm QNN, and delivering deployable bundles that meet our accuracy and latency requirements on every target platform.
Each stage in the pipeline — capture, decomposition, precision assignment, legalisation, partitioning — can affect accuracy, latency, or whether a vendor backend accepts the graph. Your work spans the full lowering stack, building compiler passes and infrastructure that scale across architectures and target platforms.
Key responsibilities
- Own the ML compilation pipeline end-to-end — from checkpoint to deployable bundle on NVIDIA (TensorRT) and Qualcomm (QNN) targets.
- Design and implement compiler passes with accuracy and latency gates, so bad compiles are caught before they reach hardware.
- Build compilation infrastructure that scales across platforms, model architectures, and SoCs — without re-engineering for each new target.
- Partner with model and training teams on compilability; build regression and benchmarking to validate changes across releases.
- Set technical direction and raise the bar for compiler engineering across the team.
Reasons to use Rodeo
I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?
Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.
Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.
Start with a chat, not a search bar
Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.
Graduate Consultant — 2026 Scheme
Why you're a good match
StrongYour economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.
See breakdownIt searches the market for you
Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.
Why you're a good match
You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.
Experience fit
Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.
Only hits
No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.
About you
- You have built or significantly extended ML compilation or graph-lowering pipelines.
- You understand multi-stage lowering (capture, decomposition, precision assignment, legalisation) and can debug what breaks at each stage.
- Strong proficiency with at least one relevant stack (e.g. MLIR, ONNX, TensorRT, Qualcomm QNN, PyTorch export/capture) and confidence learning adjacent frameworks quickly.
- Experience with quantisation in compilation — precision typing, PTQ integration, and tracking down accuracy loss from compiler transforms.
- Comfortable from high-level model graphs down to vendor backend constraints; strong Python, with C++ a plus.
- Clear communicator who can align cross-functional teams on compilation trade-offs.
- Real compiler ownership — full lowering pipeline from checkpoint to deployable bundle, working deeply with TensorRT and QNN.
- Hard problems — quantisation preservation through decomposition, cross-SoC precision typing, graph partitioning under speed/accuracy trade-offs, legalisation that does not silently break earlier passes.
- Vehicle impact — compiler passes determine what runs on embedded hardware in Wayve's driving product.
- Greenfield at Staff level — small team, high leverage, shaping the compilation stack from early stages.
- Scalable infrastructure — building pipelines that work across platforms and architectures without starting from scratch each time.


Get help with your application
Your very own career expert that helps elevate your application to the next level.
Day-to-day / scope of the role
- Own the ML compilation pipeline end-to-end on NVIDIA (TensorRT) and Qualcomm (QNN) targets.
- Design and implement compiler passes with accuracy and latency gates.
- Extend precision typing and graph-splitting logic for new architectures and SoCs.
- Partner with model and training teams on compilability.
- Build regression and benchmarking to validate changes across releases.
- Set technical direction and mentor on compiler design.
Top hard requirements (skills/experience)
- Built or owned significant parts of an ML compilation or graph-lowering pipeline.
- Deep experience with quantisation in compilation — precision typing, PTQ integration, debugging accuracy loss from compiler transforms.
- Strong Python; comfortable building and testing compiler infrastructure in production codebases.
- Proficiency with at least one of: MLIR, ONNX, TensorRT, Qualcomm QNN, PyTorch graph capture/export.
- Experience with multi-target compilation or graph partitioning across hardware backends.
- Ability to reason about correctness and performance trade-offs at each compiler stage.
“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”
Jessica, London
Location