Senior Machine Learning Research Engineer

relationrx
London, UK
2 days ago
Apply on www.adzuna.co.uk
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
£78,999.0
Working hours
Regular working hours

Tech stack

Artificial Neural Networks Cloud Computing Profiling Nvidia CUDA Distributed Computing Environment Python (Programming Language) Linux Kernel Machine Learning Open Source Technology Tensorflow Pytorch Large Language Models
+2 more
Information Technology Machine Learning Operations

Job description

We are scaling rapidly and building a team of exceptional individuals to push the boundaries of drug discovery. You will work in highly interdisciplinary teams where biology, computation, and engineering come together to solve complex problems that have not been solved before. Our state-of-the-art wet and dry labs in the heart of London are designed to accelerate this integration and translate insight into impact., * Implement and optimise large models, partnering with ML Scientists to translate research ideas into reproducible, scalable training pipelines.

  • Profile and optimise training across compute, memory, and I/O, pursuing measurable gains in throughput, convergence, and stability.
  • Design and implement distributed training strategies across multi-GPU and multi-node configurations.
  • Build and maintain core ML infrastructure.
  • Contribute to architectural and algorithmic decisions, bringing engineering judgment into research discussions.
  • Optimise inference and downstream deployment so models can be used by data scientists and biologists in our discovery workflows.
  • Address numerical, performance, and reliability issues across the stack.
  • Establish and maintain engineering practices in research code.
  • Track developments in ML systems and bring relevant advances into our stack.

Requirements

  • A degree in Computer Science, Engineering, Physics, or a related quantitative discipline; industry experience as an ML / research engineer working on large neural network training.
  • Strong software engineering fundamentals in Python and deep expertise in PyTorch (or equivalent modern ML frameworks).
  • Hands-on experience training large neural networks at scale, including distributed training frameworks.
  • Demonstrable experience profiling and optimising GPU workloads.
  • Working knowledge of cloud-based ML infrastructure and containerised environments.
  • A track record of taking research code from prototype to robust, reusable infrastructure that other people actually use. Bonus experience:

  • CUDA / Triton kernel development; FlashAttention-style attention implementations; experience with foundation models for biology, vision, or language; contributions to open-source ML frameworks.

  • Are comfortable working in a matrixed environment, balancing multiple stakeholders and contributing effectively across teams.
  • Take ownership of your work, proactively seek opportunities to contribute, and enable others to do their best work.
  • Communicate openly and directly, give and receive feedback constructively, and handle challenging conversations with respect.
  • Actively seek out diverse perspectives, build strong working relationships, and contribute to shared goals across teams.
  • Embrace challenges with openness and resilience, set high standards for yourself, and strive to deliver meaningful outcomes.

At Relation, we operate in a matrixed, interdisciplinary environment, where impact is driven through collaboration across scientific, technical, and operational domains. We collaborate, and you will partner with colleagues across multiple teams and projects, contributing your expertise while aligning to shared company priorities. We work together and win together! The patient is waiting!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.co.uk
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · World Congress 2026 Europe

1:39 min

Fundamentals of tensors and the TensorFlow library

Håkan Silfvernagel · LIVE

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

2:15 min

Open-source community and machine learning frameworks

Gian Marco Iodice Gian Marco Iodice · World Congress 2025

Videos

See all

Related articles

See all