Principal Data Engineer

Raytheon
Harlow, UK
2 days ago
Apply on www.adzuna.co.uk
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
£50,000.0 - £90,000.0
Working hours
Regular working hours

Tech stack

Airflow Amazon Web Services Amazon S3 Microsoft Azure Cloud Computing Extract Transform Load (ETL) Distributed Computing Environment Hadoop Distributed File System Python (Programming Language) Microsoft Message Queuing NoSQL OpenShift
+16 more
Cloud Services Standard Sql Unstructured Data Workflow Management Systems Software Repository Apache Spark Pandas Containerization Data Lakes Pyspark Git Flow Kubernetes Apache Kafka Apache Nifi Video Streaming Docker

Job description

  • Build and maintain data processing pipelines that clean, transform, and aggregate data from disparate sources.
  • Transform and optimise data for analytical use.
  • Collaborate with stakeholders and other engineers.
  • Contribute to the completion of project milestones.
  • Contribute to continuous improvement within the team.
  • Collaborate with peers on the teams technical direction.

Technologies:

  • Airflow
  • AWS
  • Azure
  • Cloud
  • Docker
  • ETL
  • GCP
  • Kafka
  • Kubernetes
  • NoSQL
  • OpenShift
  • Python
  • PySpark
  • SQL
  • Security
  • Spark
  • pandas
  • Support

Requirements

  • We require candidates to hold active eDV clearance.
  • Experience working in an Agile delivery team.
  • Strong analytical skills related to working with unstructured datasets.
  • Knowledge of Python, including PySpark, Pandas, and PyArrow.
  • Knowledge of distributed data processing, such as Apache Spark.
  • Knowledge of data ETL tools, such as Apache Airflow, AWS Step Functions, or Apache NiFi.
  • Experience with cloud services such as AWS, Azure, or GCP.
  • Knowledge of messaging and streaming technologies such as Kafka, AWS SQS, or other cloud queuing services.
  • Knowledge of SQL and NoSQL storage, such as HDFS, Iceberg, Elastic, S3, or data lakes.
  • Knowledge of containerisation and orchestration tools such as Docker, Kubernetes, or Openshift.
  • Knowledge of testing frameworks and best practices.
  • Familiarity with code repositories, branching strategies, pull requests, and merge processes.
  • We can provide training and development in some areas; candidates do not need expertise in every listed area.

Benefits & conditions

We are Raytheon UK, a defence and aerospace technology company working across the UK and supporting government and international customers with solutions across land, sea, air, space, and cyberspace, as well as digital and training transformation. Our National Security Cyber business delivers mission-critical solutions in a mature, agile environment, where SCRUM teams work with customers on complex digital challenges. This is a permanent, onsite Principal Data Engineer role based in London, within our experienced Software Engineering function and a cross-functional Agile team. We offer a competitive salary, a contributory pension with up to 10.5% company contribution, life assurance, 25 days holiday increasing with service, public holidays, the option to buy or sell up to five days, a discretionary bonus, flexible benefits, enhanced sick pay, and enhanced family-friendly leave. Our 37-hour working week includes an early Friday finish; working arrangements may vary by role and site.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.co.uk
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:33 min

Refactoring data science workflows using Rapids QDF and Pandas

Paul Graham Paul Graham · LIVE

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all