Performance Engineer (Junior) AI Infrastructure Cambridge (Hybrid)
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+3 more
Requirements
- A postgraduate research background, ideally a PhD in computer science, mathematics, physics or a closely related field, strongly preferred, exceptional recent master’s graduates with directly relevant coursework will also be considered
- A genuine, demonstrated grasp of computer architecture fundamentals and how LLMs and deep learning models actually run on hardware, training versus inference, matrix multiplication, KV-caching
- Real experience building performance models or forecasting tools, Python or spreadsheet-based, from research, a thesis, a placement, or serious personal projects
- Hands-on work with GPU or accelerator code, CUDA or similar
- Familiarity with profiling tools (Nsight, PyTorch Profiler) and ideally some exposure to monitoring stacks (Prometheus, Grafana)
- Strong Python for data work, Pandas and NumPy, genuine scripting ability
Nice to have: exposure to inference serving frameworks like vLLM, published research, or open source contributions in this space.
Benefits & conditions
This particular seat is for an early career professional. We’re after someone academically exceptional, ideally with a PhD, who’s ready to get stuck into real technical work quickly rather than needing a long runway to get there. Day to day You’d work alongside senior engineers on the team, pulling real metrics off live training and inference jobs and turning them into models and calculators that answer actual questions, whether an optimisation is worth shipping, whether a different setup would run cheaper. Real ownership early, with senior support close by.
About the company
Performance Engineer (Junior) AI Infrastructure Cambridge (Hybrid) up to 70k Most engineers find out whether a change works after it ships. This role is about knowing before anyone spends a penny on new hardware, building the models that predict itMy client is a Cambridge-based non-profit built on a fairly simple premise: different bits of the AI world keep solving the same infrastructure problems separately, and that’s wasteful. So they’ve built a shared space where startups, big enterprises, government bodies and university researchers can pool that hard technical work instead. Early days as an organisation, but real backing and real momentum behind it.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Data Engineer Salary UK
What Are Large Language Models?
Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud
Software Engineer Salary London