AI Systems Research Engineer

microTECH Global Limited
Edinburgh, UK
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence C++ (Programming Language) Profiling Distributed Systems Fault Tolerance Systems Theories Python (Programming Language) Machine Learning Performance Tuning AI Infrastructure Load Balancing Large Language Models
+3 more
Information Technology Machine Learning Operations TensorRT

Job description

We are seeking Systems Research Engineers with a strong interest in computer systems, distributed AI infrastructure, and performance optimization. These roles are ideal for recent PhD graduates or exceptional BSc/MSc engineers looking to build research-driven engineering experience in areas such as operating systems, distributed systems, AI model serving, and machine learning infrastructure. You will work closely with senior architects on real-world projects, helping to prototype and optimize next-generation AI infrastructure.

Requirements

Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, or related field.

Strong knowledge of distributed systems, operating systems, machine learning systems architecture, Inference serving, and AI Infrastructure.

Hands-on experience with LLM serving frameworks (e.g., vLLM, Ray Serve, TensorRT-LLM, TGI) and distributed KV cache optimization.

Proficiency in C/C++, with additional experience in Python for research prototyping.

Solid grounding in systems research methodology, distributed algorithms, and profiling tools.

Team-oriented mindset with effective technical communication skills.

Desired Qualifications and Experience:

PhD in systems, distributed computing, or large-scale AI infrastructure.

Publications in top-tier systems or ML conferences (NSDI, OSDI, EuroSys, SoCC, MLSys, NeurIPS, ICML, ICLR).

Understanding of load balancing, state management, fault tolerance, and resource scheduling in large-scale AI inference clusters.

Prior experience designing, deploying, and profiling high-performance cloud or AI infrastructure systems.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on find.jobs

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:42 min

Dissecting artificial intelligence layers from compute to applications

Christian Nagel Christian Nagel +3 · World Congress 2026 Europe

5:25 min

Implementing redundancy, failover, and architectural load balancing patterns

Mihaela-Roxana Ghidersa · LIVE

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · World Congress 2026 Europe

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · World Congress 2025

4:35 min

Learning resources and community engagement for AI engineers

Alfonso Graziano Alfonso Graziano · Coffee With Developers

12:17 min

Managing edge cases and load balancing SSE

Rainer Stropek Rainer Stropek · LIVE

Videos

See all

Related articles

See all