Senior Machine Learning Engineer

Digital Talent Agency
Málaga, Spain
11 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Training Data Artificial Intelligence Distributed Computing Environment Machine Learning Raw Data Pytorch Large Language Models Deep Learning Backend Machine Learning Operations TensorRT Data Pipelines

Job description

Most AI products are wrappers. We’re building the real thing, an agent that takes on genuine tasks for everyday users: running errands, managing workflows, holding context across long and complex conversations. Reliable by design, not by luck.We’re small, we move fast, and the ML layer is the product. We need someone to own it.The roleYou’ll bridge research and production, taking ideas and turning them into systems that run at scale, stay reliable, and get better over time. Full-stack ML ownership: from raw data to deployed model.Day to day that looks like:Building end-to-end pipelines across data, training, evaluation, and inferenceAdapting and fine-tuning models with modern techniques: LoRA, QLoRA, SFT, DPO, distillationArchitecting inference systems that hold up under real latency and cost constraintsCreating data pipelines that produce high-quality synthetic and real-world training dataRunning evaluation that goes beyond benchmarks: robustness, safety, bias, production behaviourOwning deployment: GPU optimisation, quantisation, memory efficiency, scalingWorking directly with application engineers so ML integrates cleanly into backend, mobile, and desktopYour skills and experienceDeep understanding of deep learning and transformer architecturesProven experience training, fine-tuning, or shipping large-scale models in productionStrong with at least one major ML framework (PyTorch, JAX) and quick to pick up othersFamiliar with distributed training and inference tooling: DeepSpeed, FSDP, Megatron, ZeRO, RayEngineering discipline: code that’s readable, robust, and maintainableExperience optimising for GPU constraints: quantisation, mixed precision, memoryComfortable taking ownership of ambiguous problems from zero to oneShips, iterates, learns from productionNice to haveLLM inference frameworks: vLLM, TensorRT-LLM, FasterTransformerRLHF: PPO, DPO, ORPOOpen-source contributions to ML or systems librariesScientific computing, compiler, or GPU kernel experienceAt a big company, ML work gets absorbed into a machine. Here, your systems are the product. You’ll work closely with research and engineering leadership, have real influence over how the architecture evolves, and see the direct impact of your work on users. If you want to build ML infrastructure that actually matters, this is it.#J-*****-Ljbffr

Requirements

Proven experience training, fine-tuning, or shipping large-scale models in production Strong with at least one major ML framework (PyTorch, JAX) and quick to pick up others Familiar with distributed training and inference tooling: DeepSpeed, FSDP, Megatron, ZeRO, Ray Engineering discipline: code that’s readable, robust, and maintainable Experience optimising for GPU constraints: quantisation, mixed precision, memory Comfortable taking ownership of ambiguous problems from zero to one Ships, iterates, learns from production

About the company

LLM inference frameworks: vLLM, TensorRT-LLM, FasterTransformer RLHF: PPO, DPO, ORPO Open-source contributions to ML or systems libraries Scientific computing, compiler, or GPU kernel experience At a big company, ML work gets absorbed into a machine. Here, your systems are the product. You’ll work closely with research and engineering leadership, have real influence over how the architecture evolves, and see the direct impact of your work on users. If you want to build ML infrastructure that actually matters, this is it. #J-*****-Ljbffr

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

4:09 min

Challenges of interpreting raw data with language models

Clemens Vasters Clemens Vasters · WWC 2025

4:33 min

Core stages of training and managing machine learning models

Boris Krumrey +2 · LIVE

4:41 min

Replacing PyTorch with ONNX runtime for AWS Lambda deployments

Marek Suppa · LIVE

Videos

See all

Related articles

See all