AI Platform Engineer - LLM Infrastructure

Talenzon group
London, UK
1 day ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Computer Programming Monitoring of Systems Python (Programming Language) Search Technologies Web Services Google Cloud Large Language Models Generative AI
+6 more
Backend Containerization AI Platforms Kubernetes Machine Learning Operations Docker

Job description

AI Platform Engineer - LLM Infrastructure,April 13, 2026### Job DescriptionLocation: London, UK Work Model: On-site Role Type: Full-TimeWe are looking for an AI Platform Engineer with strong experience in LLM infrastructure and scalable AI systems to join our client’s on-site team in London.This role focuses on building and maintaining internal platforms that power Generative AI applications, enabling scalable model deployment, inference, and experimentation.—### What You’ll Do* Build and maintain infrastructure for LLM deployment and inference* Develop scalable systems for embeddings, vector search, and RAG pipelines* Design APIs and services for AI model consumption* Optimise performance, latency, and cost of AI workloads* Collaborate with AI engineers and data scientists to productionise models* Implement monitoring and observability for AI systems* Support experimentation and model lifecycle management—### What We’re Looking For#### **Required Skills &

Experience* Strong experience with cloud platforms such as Amazon Web Services, Google Cloud, or Microsoft Azure* Experience with containerisation using Docker and orchestration via Kubernetes* Experience working with LLMs, embeddings, and vector databases* Strong programming skills (Python preferred)* Experience designing scalable backend systems—#### **Nice to Have* Experience with RAG architectures and GenAI frameworks* Familiarity with model serving frameworks and inference optimisation* Knowledge of MLOps workflows—Location: London, UK Work Model: On-site Role Type: Full-TimeLocation,Experience levelMid-Senior level## Work Location

Requirements

Experience* Strong experience with cloud platforms such as Amazon Web Services, Google Cloud, or Microsoft Azure* Experience with containerisation using Docker and orchestration via Kubernetes* Experience working with LLMs, embeddings, and vector databases* Strong programming skills (Python preferred)* Experience designing scalable backend systems—#### **Nice to Have* Experience with RAG architectures and GenAI frameworks* Familiarity with model serving frameworks and inference optimisation* Knowledge of MLOps workflows—Location: London, UK Work Model: On-site Role Type: Full-TimeLocation,Experience levelMid-Senior level## Work Location #J-18808-Ljbffr

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

4:57 min

Centralizing LLMOps workflows within Azure AI Foundry

Maxim Salnikov Maxim Salnikov · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all