Senior ML Performance Engineer - Real-Time Inference & Scale

Odyssey
Greater London, UK
19 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours

Tech stack

Machine Learning Performance Tuning Software Engineering Pytorch

Requirements

Odyssey in Greater London seeks an experienced software engineer specializing in machine learning performance optimization. You will optimize models for real-time users, design distributed training strategies, and work with elite ML researchers. Candidates should have at least 8 years of software engineering experience, deep insights into machine learning architectures, and proficiency in PyTorch and NVIDIA optimization. This position offers autonomy in technical decisions and a chance to work with cutting-edge technology. #J-18808-Ljbffr

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

8:32 min

Benchmarking GitOps engine constraints for extensive multi-cluster environments

Artem Lajko · Europe 2026 Virtual

3:30 min

Optimizing performance using dedicated open source inference engines

Patrick Koss Patrick Koss · World Congress 2025

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

57 sec

Implementing predictive prefetch with machine learning

Jessica Janiuk · JS Congress

Videos

See all

Related articles

See all