Contract Data Engineer - NLP, LLM

Future plc
London, UK
2 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Airflow Amazon Web Services Microsoft Azure Information Engineering Data Infrastructure Extract Transform Load (ETL) Python (Programming Language) Natural Language Processing NoSQL Performance Tuning Standard Sql Unstructured Data
+7 more
Large Language Models Prompt Engineering Pyspark Information Technology Apache Kafka Machine Learning Operations Data Pipelines

Job description

This range is provided by Future Talent Group. Your actual pay will be based on your skills and experience - talk with your recruiter to learn more.

Overview

We are seeking a highly skilled Contract Data Engineer with proven expertise in Natural Language Processing (NLP) and Large Language Models (LLMs). The ideal candidate will be responsible for designing, building, and optimizing data pipelines and infrastructure to support NLP/LLM-driven applications and insights. You’ll work closely with our data science and ML teams to enable robust, scalable, and production-ready solutions.

Responsibilities

  • Design, develop, and maintain scalable data pipelines for NLP and LLM workloads.
  • Build and optimize data infrastructure to support training, fine-tuning, and deployment of LLM models.
  • Work with cloud platforms (AWS, GCP, or Azure) to manage data and ML infrastructure.
  • Develop and maintain ETL/ELT processes for structured and unstructured data.

Qualifications

  • Experience working with Large Language Models (LLMs), including fine-tuning, prompt engineering, and integration.
  • Proficiency in Python and common data engineering frameworks (e.g., PySpark, Airflow, dbt, Kafka).
  • Strong knowledge of SQL and relational as well as NoSQL databases.

Employment type

  • Contract

Job function

  • Information Technology

Industries

  • Staffing and Recruiting

We are not including location-based or time-based postings here. If you are interested, please contact Future Talent Group for more details about the role.

Requirements

  • Experience working with Large Language Models (LLMs), including fine-tuning, prompt engineering, and integration.
  • Proficiency in Python and common data engineering frameworks (e.g., PySpark, Airflow, dbt, Kafka).
  • Strong knowledge of SQL and relational as well as NoSQL databases.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Applying large language models to infrastructure tasks

Alfonso Sandoval Rosas Alfonso Sandoval Rosas · Europe 2026 Virtual

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all