Principal AI Data Engineer

Certain Advantage
London, UK
1 day ago

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Artificial Intelligence Automation of Tests Microsoft Azure Computer Programming Data Architecture Information Engineering Extract Transform Load (ETL) DevOps Github Python (Programming Language) SQL Databases Azure Data Factory
+11 more
Large Language Models Prompt Engineering Generative AI Data Lakes Qlikview Spark Streaming Machine Learning Operations Virtual Agents GPT Data Pipelines Databricks

Job description

This role is an initial contract until the end of the year and is hybrid, requiring 3 days per week onsite in London.

You’ll be responsible for:

  • Developing AI, GenAI and Agentic AI prototypes using Azure AI Foundry, Copilot Studio, Databricks Mosaic AI and MLflow.
  • Building, optimising and evaluating Retrieval-Augmented Generation (RAG) solutions using prompt engineering and embedding models.
  • Designing and deploying AI agents using frameworks including LangChain, LangGraph and AutoGen.
  • Delivering production-ready GenAI applications for commercial teams within Trading & Supply.
  • Building scalable AI workflows on Databricks using Mosaic AI, MLflow, Genie and AgentBricks.
  • Applying DevOps best practices through CI/CD pipelines, GitHub workflows and automated testing.
  • Designing evaluation frameworks to benchmark, test and improve GenAI system performance and reliability.

Requirements

  • Strong expertise with Databricks, including Delta Lake, Unity Catalog, DLT and MLflow.
  • Deep knowledge of Large Language Models including GPT, Llama, Claude and Mistral.
  • Experience building and deploying GenAI, RAG and Agentic AI solutions in enterprise environments.
  • Strong programming skills in Python, SQL and/or Scala.
  • Experience with Azure cloud technologies including Azure OpenAI and Azure AI Foundry.
  • Knowledge of Spark Structured Streaming, Autoloader and modern data engineering practices.
  • Experience with ETL/ELT, data modelling and scalable data architectures.
  • Familiarity with Azure Data Factory, Qlik Replicate and data ingestion pipelines.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.totaljobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · World Congress 2024

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · World Congress 2025

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

Videos

See all

Related articles

See all