Performance Engineer (Junior) AI Infrastructure Cambridge (Hybrid)

Pure Resourcing Solutions
Dry Drayton, UK
4 days ago
Apply on find.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Starter
Compensation
£55,000.0 - £70,000.0
Working hours
Regular working hours
Job source

Tech stack

Spreadsheets Profiling Nvidia CUDA Python (Programming Language) NumPy Open Source Technology Prometheus AI Infrastructure Pytorch Large Language Models Grafana Deep Learning
+3 more
Caching Pandas Information Technology

Requirements

  • A postgraduate research background, ideally a PhD in computer science, mathematics, physics or a closely related field, strongly preferred, exceptional recent master’s graduates with directly relevant coursework will also be considered
  • A genuine, demonstrated grasp of computer architecture fundamentals and how LLMs and deep learning models actually run on hardware, training versus inference, matrix multiplication, KV-caching
  • Real experience building performance models or forecasting tools, Python or spreadsheet-based, from research, a thesis, a placement, or serious personal projects
  • Hands-on work with GPU or accelerator code, CUDA or similar
  • Familiarity with profiling tools (Nsight, PyTorch Profiler) and ideally some exposure to monitoring stacks (Prometheus, Grafana)
  • Strong Python for data work, Pandas and NumPy, genuine scripting ability

Nice to have: exposure to inference serving frameworks like vLLM, published research, or open source contributions in this space.

Benefits & conditions

This particular seat is for an early career professional. We’re after someone academically exceptional, ideally with a PhD, who’s ready to get stuck into real technical work quickly rather than needing a long runway to get there. Day to day You’d work alongside senior engineers on the team, pulling real metrics off live training and inference jobs and turning them into models and calculators that answer actual questions, whether an optimisation is worth shipping, whether a different setup would run cheaper. Real ownership early, with senior support close by.

About the company

Performance Engineer (Junior) AI Infrastructure Cambridge (Hybrid) up to 70k Most engineers find out whether a change works after it ships. This role is about knowing before anyone spends a penny on new hardware, building the models that predict itMy client is a Cambridge-based non-profit built on a fairly simple premise: different bits of the AI world keep solving the same infrastructure problems separately, and that’s wasteful. So they’ve built a shared space where startups, big enterprises, government bodies and university researchers can pool that hard technical work instead. Early days as an organisation, but real backing and real momentum behind it.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on find.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:37 min

Optimizing technical profiles for AI sourcing and recruitment

Mina Golesorkhi Mina Golesorkhi · World Congress 2026 Europe

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham · World Congress 2025

Videos

See all

Related articles

See all