> Markdown version of [/jobs/ext/2970773-performance-engineer-junior-ai-infrastructure-cambridge-hybrid](https://www.wearedevelopers.com/jobs/ext/2970773-performance-engineer-junior-ai-infrastructure-cambridge-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Performance Engineer (Junior) AI Infrastructure Cambridge (Hybrid) - **Company:** Pure Resourcing Solutions - **Location:** Dry Drayton, UK - **Experience:** Starter - **Salary:** £55,000.0 - £70,000.0 - **Contract:** Permanent contract - **Skills:** Spreadsheets, Profiling, Nvidia CUDA, Python (Programming Language), NumPy, Open Source Technology, Prometheus, AI Infrastructure, Pytorch, Large Language Models, Grafana, Deep Learning, Caching, Pandas, Information Technology - **Published:** September 18, 2026 - **Apply:** https://find.jobs/jobs-near-me/apply/ats-redirect/?id=2976499002-2 ## About the Role * A postgraduate research background, ideally a PhD in computer science, mathematics, physics or a closely related field, strongly preferred, exceptional recent master's graduates with directly relevant coursework will also be considered * A genuine, demonstrated grasp of computer architecture fundamentals and how LLMs and deep learning models actually run on hardware, training versus inference, matrix multiplication, KV-caching * Real experience building performance models or forecasting tools, Python or spreadsheet-based, from research, a thesis, a placement, or serious personal projects * Hands-on work with GPU or accelerator code, CUDA or similar * Familiarity with profiling tools (Nsight, PyTorch Profiler) and ideally some exposure to monitoring stacks (Prometheus, Grafana) * Strong Python for data work, Pandas and NumPy, genuine scripting ability Nice to have: exposure to inference serving frameworks like vLLM, published research, or open source contributions in this space. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [Serverless deployment of (large) NLP models ](https://www.wearedevelopers.com/videos/158-serverless-deployment-of-large-nlp-models) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development)