Lead AI Engineer (Gen AI Platform, Agentic AI & LLM Infrastructure & Orchestration)

G-Research
New Addington, UK
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Cluster Analysis Continuous Integration Distributed Systems Machine Learning Open Source Technology Prometheus Systems Integration AI Infrastructure Datadog System Availability
+10 more
Delivery Pipeline Large Language Models Grafana Kubernetes Information Technology Data Management Machine Learning Operations Terraform Docker Service Stack

Job description

We tackle the most complex problems in quantitative finance, by bringing scientific clarity to financial complexity.

From our London HQ, we unite world-class researchers and engineers in an environment that values deep exploration and methodical execution - because the best ideas take time to evolve. Together we’re building a world-class platform to amplify our teams’ most powerful ideas.

As part of our engineering team, you’ll shape the platforms and tools that drive high-impact research - designing systems that scale, accelerate discovery and support innovation across the firm.

Take the next step in your career.

The role

The Core AI team is a centralised infrastructure team within the AI Engineering department. We build, operate and scale the foundational platform that powers AI innovation across G-Research, including on-prem open model inference, model serving, AI developer experience tooling, centralised MCP servers and secure agent sandboxing.

We provide the foundations that enable teams across the firm to innovate and deliver with confidence, working closely with our Applied AI team.

As an Engineer in Core AI, you will work across four key areas:

Infrastructure and serving - design, build and operate on-prem model inference and serving platformsMCP server infrastructure - build and operate centralised MCP servers that provide secure, governed access to tools and dataSecurity and sandboxing - design and implement infrastructure for safe execution of autonomous AI agents in a regulated environmentDeveloper AI experience - improve developer experience through seamless integrations and user-facing tools., Designing and operating model serving infrastructure, including inference pipelines and scheduling systemsBuilding and running centralised MCP servers, ensuring secure, reliable access to enterprise tools and dataOwning platform reliability, performance and scalability across Kubernetes-based infrastructure, including observability, capacity planning and incident responseBuilding self-service tooling and APIs to enable teams to provision and consume AI infrastructure independentlyIntegrating platform services with existing technology stacks, ensuring clear interfaces, monitoring and CI/CDEvaluating and adopting open-source technologies and applying emerging best practices to improve the platform

Requirements

We value pragmatic engineers who combine deep infrastructure expertise with strong systems thinking and clear communication. You should enjoy building reliable, secure platforms at scale - the kind of foundations that hundreds of engineers and quants depend on daily without needing to think about.

The ideal candidate will have the following skills and experience:

Essential:

Strong expertise in C# and Python, building distributed systems and platform-level softwareDeep Kubernetes expertise, including multi-tenant cluster operations and platform extensionsExperience with Docker, Terraform and CI/CD in controlled or regulated environmentsStrong understanding of distributed systems, including networking, storage, security and performanceExperience with model serving and inference infrastructure, including deployment, scaling and optimisation of open modelsClear communication skills, with ability to explain complex concepts and produce high-quality technical documentation

Desirable:

Experience with MCP or similar platform servicesFamiliarity with sandboxing and workload isolation technologiesExperience in quantitative finance or low-latency systemsAWS experience particularly in hybrid environmentsExperience with observability tooling such as Prometheus, Grafana or OpenTelemetryContributions to open-source projects in relevant domains, Senior (5+ years of experience)

Tagged as: Clustering, Industry, NLP, United Kingdom

Benefits & conditions

Why join us?Highly competitive compensation plus annual discretionary bonusLunch provided (via Just Eat for Business) and dedicated barista bar30 days annual leave9% company pension contributionsInformal dress code and excellent work/life balanceComprehensive healthcare and life assuranceCycle-to-work schemeMonthly company events.

About the company

G-Research, G-Research is a research and technology company that specializes in Quantitative Research and IT Infrastructure.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on nlppeople.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

1:36 min

Visualizing memory limits and isolating suspicious endpoints

Dina Matveev Dina Matveev · Europe 2026 Virtual

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

3:22 min

Evaluating advanced artificial intelligence platforms for daily recruitment

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

Videos

See all

Related articles

See all