> Markdown version of [/jobs/ext/2022209-associate-data-engineer](https://www.wearedevelopers.com/jobs/ext/2022209-associate-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Associate Data Engineer - **Company:** Quantexa - **Location:** London, UK (Remote available) - **Salary:** £39,037.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Agile Methodology, Airflow, Data Analysis, Big Data, Software as a Service, Code Reuse, Extract Transform Load (ETL), Elasticsearch, Python (Programming Language), Parsing, Software Deployment, Data Processing, Apache Spark, Data Lakes, Data Pipelines - **Published:** August 11, 2026 - **Apply:** https://www.adzuna.co.uk/jobs/details/5834863862 ## About the Role * A strong coding background, ideally in Scala or otherwise in a relevant language that will allow you to learn Scala quickly (e.g. Java/Python) * Big data, either from a software deployment/implementation or a data science perspective * Working with big data technology, ideally Spark but others will also be useful such as Airflow or Elasticsearch * Working in an Agile environment * Building data processing pipelines for use in production batch systems, including either traditional ETL pipelines and/or analytics pipelines * Manipulating data through cleansing, parsing, standardising etc, especially in relation to improving data quality/integrity * Building and deploying SaaS products ## Description * Developing Quantexa's libraries for cleansing, parsing and standardising data used in entity resolution * Finding efficiency/performance improvements through big data testing and building performance tooling * Owning best practices in entity resolution and network building Data Feeds: * Building standardised and reusable code for processing various third party/open source data sets * Managing an internal data lake for the provision of this data by other teams for testing and analytics * Owning general best practices for ingesting and processing data to get it ready for use in the Quantexa Platform, including pipelines and scheduling Demos: * Developing, deploying and maintaining all Quantexa demos, showcasing the different use cases for the Quantexa Platform * Owning the Quantexa Trial platform, for prospective Quantexa clients to see the product in action using real data provided by Data Feeds * Building tools to enable solution owners and sales to create their own custom demos SaaS: * Building Quantexa's emerging SaaS offering, a cloud hosted, standardized deployment of the Quantexa Platform * Targeting mid-market banks in the US for Retail AML initially, providing them with a cost-effective Quantexa solution, then expanding in future to more use cases and geographies * Implementing cutting edge features of the Quantexa Platform ensuring SaaS customers always on the latest and greatest of Quantexa More detailed information on the day to day activities of the four sub-teams can be shared on request. The teams all work together closely and team members are able to rotate between the teams to enable knowledge sharing and personal development. We are looking for candidates who enjoy the following to join us: * Data processing/ETL pipelines * Analysing and examining real and varied data * Full stack development, but with a heavy focus on the data processing/ETL side * Solving difficult problems with efficient, resilient, high impact code * Working in the cloud with production-grade systems * Defining best-practices and sharing expertise you've developed * Working in a fast moving, Agile environment * Growing and thriving within one of the UK's fastest growing scale-ups ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)