> Markdown version of [/jobs/ext/2970568-performance-engineer-junior-ai-infrastructure-cambridge-hybrid](https://www.wearedevelopers.com/jobs/ext/2970568-performance-engineer-junior-ai-infrastructure-cambridge-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Performance Engineer (Junior) | AI Infrastructure | Cambridge (Hybrid) - **Company:** Itmy - **Location:** Cambridge, UK - **Experience:** Starter - **Salary:** £70,000.0 - **Contract:** Permanent contract - **Skills:** Spreadsheets, Profiling, Nvidia CUDA, Python (Programming Language), NumPy, Open Source Technology, Prometheus, AI Infrastructure, Pytorch, Large Language Models, Grafana, Deep Learning, Pandas, Information Technology - **Published:** September 18, 2026 - **Apply:** https://itjobpro.co.uk/job/performance-engineer-junior-ai-infrastructure-cambridge-hybrid ## About the Role A postgraduate research background, ideally a PhD in computer science, mathematics, physics or a closely related field, strongly preferred, exceptional recent master's graduates with directly relevant coursework will also be considered A genuine, demonstrated grasp of computer architecture fundamentals and how LLMs and deep learning models actually run on hardware, training versus inference, matrix multiplication, KV-caching Real experience building performance models or forecasting tools, Python or spreadsheet-based, from research, a thesis, a placement, or serious personal projects Hands-on work with GPU or accelerator code, CUDA or similar Familiarity with profiling tools (Nsight, PyTorch Profiler) and ideally some exposure to monitoring stacks (Prometheus, Grafana) Strong Python for data work, Pandas and NumPy, genuine scripting abilityNice to have: exposure to inference serving frameworks like vLLM, published research, or open source contributions in this space. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/859-accelerating-python-on-gpus) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/1112-accelerating-python-on-gpus) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)