> Markdown version of [/jobs/ext/2970570-senior-performance-engineer-ai-infrastructure-cambridge-hybrid](https://www.wearedevelopers.com/jobs/ext/2970570-senior-performance-engineer-ai-infrastructure-cambridge-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Performance Engineer | AI Infrastructure | Cambridge (Hybrid) | - **Company:** Pure Resourcing Solutions - **Location:** Cambridge, UK - **Experience:** Expert - **Salary:** £90,000.0 - £120,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Apache HTTP Server, Spreadsheets, Profiling, Nvidia CUDA, Python (Programming Language), NumPy, Open Source Technology, Reliability Engineering, Prometheus, AI Infrastructure, Pytorch, Large Language Models, Grafana, Deep Learning, Pandas, Information Technology - **Published:** September 18, 2026 - **Apply:** https://www.reed.co.uk/jobs/senior-performance-engineer--ai-infrastructure--cambridge-hybrid-/57359532 ## About the Role * A degree in computer science, mathematics, or something adjacent * A track record of building performance models or calculators (Python or spreadsheet-based) that actually forecast how a system will behave * Hands-on GPU/accelerator code optimisation, CUDA or similar * Genuine understanding of how LLMs and deep learning models run on real hardware, training versus inference, matrix multiplication, KV-caching, that level of detail * Comfortable in profiling tools like Nsight or PyTorch Profiler, and monitoring stacks like Prometheus and Grafana * Python for data work, Pandas and NumPy, plus general scripting Nice to have rather than essential: a postgraduate degree and research background (publications welcome), real depth on inference serving frameworks like vLLM, a stats background, and any open source or research contributions., * Analytical Thinking (Solving) * Apache Web Server * Communication Skills (Key) * Hewlett Packard (HP) Products * Performance Engineering * Performance Management Roles (System) * Problem Solving (Process) * Python Programming (Beginner) * Reliability Engineering * Stakeholder Management (Business) ## Description Senior Performance Engineer | AI Infrastructure | Cambridge (Hybrid) | £90k-£120k Nobody quite knows where their compute budget is actually going until someone builds the model that tells them. That's this role. My client is a Cambridge-based non-profit that exists to stop different parts of the AI world quietly rebuilding the same infrastructure. Rather than a startup, a big enterprise, a government department and a university lab each working out GPU efficiency from scratch, they pool the hard problems and the expertise needed to solve them, so everyone moves faster. It's early days for the organisation but there's serious momentum and serious financial backing behind it already. They're hiring Performance Engineers at junior and senior level, to sit at the sharp end of that mission. Day to day You'd sit between the research and engineering teams, pulling real numbers off live training and inference runs rather than working from theory. From there, the job is building the models and calculators that turn those numbers into an actual answer: will this optimisation help, would a different accelerator be worth the spend, is this architecture change going to pay for itself. Those answers don't stay internal either, they shape what gets bought and how systems get built, for the organisation itself and for everyone else in the membership relying on that judgement. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/859-accelerating-python-on-gpus) - [How AI Models Get Smarter](https://www.wearedevelopers.com/videos/1374-how-ai-models-get-smarter) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud)