Senior Data Engineer (eDV)

Frontier Resourcing
Cheltenham, UK
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Airflow Amazon Web Services Amazon S3 Microsoft Azure Cloud Computing Code Review Extract Transform Load (ETL) Distributed Computing Environment Elasticsearch Revision Control Systems Hadoop Distributed File System
+22 more
Python (Programming Language) Microsoft Message Queuing NoSQL OpenShift SQL Databases Unstructured Data Software Organization Data Storage Technologies Apache Spark Pandas Containerization Data Lakes Pyspark Git Flow Kubernetes Apache Kafka Apache Nifi Data Management Video Streaming Data Pipelines Serverless Computing Docker

Job description

As a Senior Data Engineer, you will join a collaborative Agile environment, playing a key role in designing, building and maintaining scalable data platforms and processing pipelines. You will work alongside software engineers, architects and stakeholders to transform and optimise data for advanced analytical and operational use., * Design, develop and maintain robust data pipelines to ingest, clean, transform and aggregate data from a variety of sources.

  • Collaborate closely with engineers, technical leads and stakeholders to deliver project objectives.
  • Support the successful delivery of project milestones within Agile teams.
  • Contribute to continuous improvement initiatives, engineering best practices and team capability development.
  • Influence technical direction and contribute to architectural and design decisions.

Requirements

  • Strong analytical skills with experience working with large-scale and unstructured datasets.
  • Python development experience, including technologies such as PySpark, Pandas and PyArrow.
  • Expertise in distributed data processing frameworks, particularly Apache Spark.
  • Experience building and maintaining ETL pipelines using technologies such as Apache Airflow, AWS Step Functions or Apache NiFi.
  • Cloud platform experience across AWS, Azure and/or GCP.
  • Knowledge of messaging and streaming technologies including Kafka, AWS SQS or equivalent cloud-native services.
  • Strong understanding of SQL and NoSQL databases.
  • Experience with modern data storage technologies including HDFS, Iceberg, Elasticsearch, S3 and Data Lake architectures.
  • Containerisation and orchestration experience using Docker, Kubernetes and/or OpenShift.
  • Knowledge of testing frameworks, engineering standards and software development best practices.
  • Experience working with source control systems, branching strategies, pull requests and code review processes.

Benefits & conditions

  • Opportunity to work on highly secure, cutting-edge programmes with real-world impact.
  • Exposure to advanced technologies across AI, machine learning, cyber security and data engineering.
  • Structured training and development support to help broaden your technical capabilities.
  • Clear progression opportunities within a growing and highly regarded engineering function.
  • Collaborative Agile environment working alongside some of the industry’s leading technical specialists., * Competitive salary and excellent benefits
  • eDV bonus on top of salary and annual bonus
  • Company bonus scheme
  • Contributory Pension Scheme (up to 10.5% company contribution)
  • 6 times salary ‘Life Assurance’ with pension
  • 25 days holiday (increasing with service) + statutory public holidays, plus opportunity to buy and sell up to 5 days

About the company

We are partnering with a leading defence and technology organisation to recruit an experienced Senior Data Engineer to join their highly specialised engineering teams across Manchester, Gloucester or London.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · WWC 2024

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:33 min

Refactoring data science workflows using Rapids QDF and Pandas

Paul Graham Paul Graham · LIVE

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all