> Markdown version of [/jobs/ext/3100678-senior-machine-learning-engineer-ai-performance-london-united-kingdom](https://www.wearedevelopers.com/jobs/ext/3100678-senior-machine-learning-engineer-ai-performance-london-united-kingdom). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Machine Learning Engineer, AI Performance London, United Kingdom - **Company:** Wayve - **Location:** London, UK - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Nvidia CUDA, Software Debugging, Machine Learning, OpenCL, Performance Tuning, Toolchain, Pytorch, Deep Learning, Low Latency, Machine Learning Operations, TensorRT - **Published:** September 27, 2026 - **Apply:** https://www.apply4u.co.uk/jobs/senior-machine-learning-engineer-ai-performance-london-united-kingdom/48667829 ## About the Role We're looking for a Senior Machine Learning Engineer to join a high-ownership team responsible for delivering production-ready model releases as our OEM engagements and release cadence accelerate. This is an applied, delivery-focused MLE role-ideal for engineers who love shipping real systems and iterating quickly. You'll work on taking models from "works in training" to "meets product constraints," partnering closely with teams downstream (e.g., inference/performance specialists) to ensure models are ready for deployment on-vehicle. As model capability grows, you'll help keep the system within tight runtime constraints using a practical model optimisation techniques (e.g., quantisation, distillation, low-rank methods) where appropriate. Key responsibilities Own end-to-end delivery of model releases, from initial requirements through training, evaluation, iteration, and final readiness for deployment. Train and iterate on PyTorch models with a strong experimental approach (hypothesis-driven iteration, ablations, clear evaluation criteria). Debug and improve model performance using strong analytical skills-identifying regressions, root-causing issues, and proposing fixes. Apply optimisation techniques (e.g., quantisation and distillation where beneficial), understanding trade-offs and when methods are appropriate. Collaborate cross-functionally with adjacent ML and performance engineering teams to hand off models, define bottlenecks, and align on optimisation priorities. Communicate clearly with stakeholders to align on delivery timelines, trade-offs, and readiness criteria. About you In order to set you up for success in this role at Wayve, we're looking for the following skills and experience: Essential Proven experience improving performance in production systems with tight constraints (latency, memory, bandwidth, power/thermal, or cost). Strong hands-on experience training and iterating on deep learning models in PyTorch (not just using high-level tooling). Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction - from high-level model behaviour down to low-level kernel/runtime execution. Familiarity with model optimisation concepts such as quantisation and/or distillation (hands-on is a strong signal, but not a strict requirement if the fundamentals are solid). Ability to reason across multiple levels of abstraction-from high-level model behaviour down to practical runtime/latency implications. Strong engineering fundamentals and collaboration skills. Desirable Experience working on models that must meet tight latency / efficiency constraints (edge, embedded, real-time, or similarly constrained production settings). Exposure to ML systems spanning training * evaluation * deployment handoff (even if you're not writing kernels day-to-day). Exposure to embedded or edge deployment of ML models, including benchmarking on real devices and handling system-level constraints. #LI-HH1 Wayve is committed to creating an inclusive interview experience. If you require any accommodations or adjustments to participate fully in our interview process, please let us know. At Wayve we're committed to creating a diverse, fair and respectful culture that is inclusive of everyone based on their unique skills and perspectives, and regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, veteran status, pregnancy or related condition (including breastfeeding) or any other basis as protected by applicable law. For more information visit Careers at Wayve. To learn more about what drives us, visit Values at Wayve For US candidates only, please visit E-Verify Notice and Participation and Right to Work DISCLAIMER: We will not ask about marriage or pregnancy, care responsibilities or disabilities in any of our job adverts or interviews. However, we do look to capture information about care responsibilities, and disabilities among other diversity information as part of an optional DEI Monitoring form to help us identify areas of improvement in our hiring process and ensure that the process is inclusive and non-discriminatory. #J-18808-Ljbffr ## Related Videos - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [Developing an AI.SDK](https://www.wearedevelopers.com/videos/198-developing-an-ai-sdk) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) - [Trends, Challenges and Best Practices for AI at the Edge](https://www.wearedevelopers.com/videos/630-trends-challenges-and-best-practices-for-ai-at-the-edge) - [Localized Open Models in Production: What Builders Need to Know](https://www.wearedevelopers.com/videos/100270-localized-open-models-in-production-what-builders-need-to-know) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)