> Markdown version of [/jobs/ext/341776-data-engineer](https://www.wearedevelopers.com/jobs/ext/341776-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Addington - **Location:** UK - **Experience:** Starter - **Contract:** Permanent contract - **Skills:** Airflow, Amazon S3, Big Data, Cloud Storage, Data Validation, Information Engineering, Extract Transform Load (ETL), Data Manipulation Languages, Data Mining, Data Warehousing, Distributed Computing Environment, Distributed Systems, MapReduce, Python (Programming Language), NumPy, Azure Data Lake, SQL Stored Procedures, SQL Databases, Technical Data Management Systems, Data Processing, Scripting, Apache Spark, Git, Pandas, Data Lakes, Pyspark, Information Technology, Star Schema, Google Bigquery, Data Pipelines - **Published:** June 14, 2026 - **Apply:** https://www.apply4u.co.uk/jobs/x/38992093/ ## About the Role ProgrammingStrong proficiency in Python for data manipulation and scripting. Familiarity with standard Python data libraries (e.g., Pandas, NumPy).DatabaseExpert-level proficiency in SQL (Structured Query Language). Experience writing complex joins, stored procedures, and performing performance tuning.Big Data ConceptsFoundational understanding of Big Data architecture (Data Lakes, Data Warehouses) and distributed processing concepts (e.g., MapReduce).ETL/ELTBasic knowledge of ETL principles and data modeling (star schema, snowflake schema).Version ControlPractical experience with Git (branching, merging, pull requests). Preferred Qualifications (A Plus)Experience with a distributed computing framework like Apache Spark (using PySpark).Familiarity with cloud data services (AWS S3/Redshift, Azure Data Lake/Synapse, or Google BigQuery/Cloud Storage).Exposure to workflow orchestration tools (Apache Airflow, Prefect, or Dagster).Master's degree in Computer Science, Engineering, Information Technology, or a related field. ## Description DATA ENGINEER (Python & SQL Focus) ( MASTER'S degree is MUST )We're looking for an enthusiastic and detail-oriented Junior Big Data Developer to join our data engineering team. This role is ideal for an early-career professional with foundational knowledge in data processing, strong proficiency in Python, and expert skills in SQL. You'll focus on building, testing, and maintaining data pipelines and ensuring data quality across our scalable Big Data platforms. Key ResponsibilitiesData Pipeline Development: Assist in the design, construction, and maintenance of robust ETL/ELT pipelines to integrate data from various sources into our data warehouse or data lake.Data Transformation with Python: Write, optimize, and maintain production-grade Python scripts to clean, transform, aggregate, and process large volumes of data.Database Interaction (SQL): Develop complex, high-performance SQL queries (DDL/DML) for data extraction, manipulation, and validation within relational and data warehousing environments.Quality Assurance: Implement data quality checks and monitoring across pipelines, identifying discrepancies and ensuring the accuracy and reliability of data.Collaboration: Work closely with Data Scientists, Data Analysts, and other Engineers to understand data requirements and translate business needs into technical data solutions.Tooling & Automation: Utilize version control tools like Git and contribute to the automation of data workflows and recurring processes.Documentation: Create and maintain technical documentation for data mappings, processes, and pipelines. ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Getting to Know Your Legacy (System) with AI-Driven Software Archeology](https://www.wearedevelopers.com/videos/1437-getting-to-know-your-legacy-system-with-ai-driven-software-archeology) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries)