Staff ML Performance Engineer (Inference Optimisation)
Before the detail, here's the challenge you'd help us solve.
We build the embodied intelligence that moves real vehicles safely, and the ecosystem a billion machines will run on in the future. Very few people in AI can say this. Every role here, whatever the team, plugs into that.
Here’s what this particular role covers.
The role
As a Staff ML Performance Engineer, you’ll play a key role in high-impact projects, optimising ML inference for edge accelerators and GPUs. The focus of this team is to run large transformer-based models efficiently on low-cost, low-power edge devices to enable Wayve’s first driving product.
You’ll help set the technical direction for turning these models into production systems that run reliably on in-vehicle compute. This is a hands-on role working across ML systems, compilers, runtimes, kernels, and embedded deployment, contributing to several early-stage, high-impact projects at Wayve.
Key responsibilities:
Profile and pinpoint bottlenecks across the full inference stack (model graph, compiler/runtime, kernel execution, memory movement) and deliver measurable improvements.
Implement and validate optimisations in compilers, runtimes, and/or kernels (e.g. operator fusion, scheduling, quantisation-aware performance, custom kernels).
Build robust benchmarking and regression testing to ensure performance improvements hold across models, devices, and software releases.
Optimise for multiple targets (e.g. NVIDIA Orin/Thor, Qualcomm) and work with teams to support these in a maintainable way
Collaborate with model developers to influence architecture and training/deployment decisions that affect on-device performance.
Contribute to technical roadmaps and tooling and help raise the standard of performance engineering across the team
About you
Essential
Proven experience improving performance in production systems with tight constraints (latency, memory, bandwidth, power/thermal, or cost).
Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly.
Comfort operating at multiple levels of abstraction — from high-level model behaviour down to low-level kernel/runtime execution.
Strong software engineering fundamentals (debugging, profiling, testing, and maintainable code).
Clear communicator and collaborative teammate; able to align multiple stakeholders on performance trade-offs and priorities.
Desirable
Exposure to embedded or edge deployment of ML models, including benchmarking on real devices and handling system-level constraints.
Experience with NVIDIA and/or Qualcomm SoCs and performance tooling.
Python and C++ proficiency.
Experience mentoring others and/or driving technical direction in a small, fast-moving team.
This is a full-time role based in our office in London. At Wayve we want the best of all worlds so we operate a hybrid working policy that combines time together in our offices and workshops to fuel innovation, culture, relationships and learning, and time spent working from home.
#LI-HH1
A quick, honest note before you apply.
Wayve is not a mature, fully-structured place with the playbook already written. Much of how we work is still being written, and if you join, you’ll help write it. That suits people who want real ownership more than people who need a settled structure from day one.
If that sounds like the kind of problem you want to spend your time on, we’d really like to hear from you.
Recommended Jobs
Calypso Programme Director (Hiring Immediately)
G MASS Consulting is supporting a global broker undertaking a major front-to-back transformation of its trading and post-trade infrastructure. The programme centres on the implementation and consolid…
Principal Implementation Consultant
Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate wha…
Product Marketing Manager
The role We are looking for an experienced Product Marketing Manager to help shape and deliver product marketing across WGSN Group, including WGSN and IWSR. This is an on-site role out of our L…
Nuclear Medicine Radiographer
Nuclear Medicine Radiographer (Band 6) Job Title: Nuclear Medicine Radiographer Location: Harrow, HA1 3UJ Band: Band 6 Contract Type: Locum Salary: £23 - £26ph About you: Are you…
.NET Developer
.NET Developer, .NET 10.0, C# 14, Agile - London (Tech stack: .NET Developer, .NET 10.0, ASP.NET Core, C# 14, Azure, Angular 21, Vue.js, TypeScript, Multithreading, RESTful, ASP.NET Core Web API, EF …
Behaviour Mentor
Behaviour Mentor ASAP/October start West London Behaviour Mentor Key Stages 3 and 4 Outstanding school Interviews and trials immediately Salary - £100 - £180 JOB DESCRIPTION Beh…
Need help installing a new broadband router and setup
I'm looking for a friendly techy to help me install a new broadband router and get everything connected. I have a wired and wireless setup at home, and I want a reliable configuration for my laptop, p…
Senior Project Manager (SC Cleared) (IT)
At CGI, you'll lead the successful delivery of complex projects that help transform critical services across government and secure environments. As a Senior Project Manager, you'll work with multidisc…
Oracle Supply Chain Analyst
Engagement Type: Support Description : Talenterprize are appointed by a global electronics manufacturing company to secure an Oracle supply chain management analyst. The supply chain manageme…
Cosmetic Oculoplastic Practice Nurse
A brilliant new job opportunity has arisen for an experienced Cosmetic Oculoplastic Practice Nurse to work in a fantastic independent private hospital next to Central London. You will be working for …