> Markdown version of [/jobs/ext/2285854-data-engineer](https://www.wearedevelopers.com/jobs/ext/2285854-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** GeekSoft Consulting - **Location:** Eindhoven, Netherlands - **Experience:** Expert - **Contract:** Temporary contract - **Skills:** Agile Methodology, Amazon Web Services, Amazon S3, Microsoft Azure, Big Data, Cloud Database, Cloud Engineering, Code Reuse, Continuous Integration, Information Engineering, Data Governance, Extract Transform Load (ETL), Data Security, DevOps, Python (Programming Language), Meta-Data Management, Performance Tuning, Scrum Methodology, Role-Based Access Control, Ansible, Software Engineering, SQL Databases, AWS Cdk, Data Lakes, Gitlab-ci, Infrastructure Automation Frameworks, Information Technology, Data Lineage, Deployment Automation, AWS Glue, AWS Data Analytics, Data Management, Machine Learning Operations, Databricks - **Published:** August 29, 2026 - **Apply:** https://www.adzuna.nl/details/5859960361 ## About the Role * 5+ years of professional data engineering experience working with Big Data in enterprise IT environments. * Collaborate with project managers, resource managers, IT teams, data scientists, and analysts to gather requirements, define project scope, and deliver reliable, model-ready datasets. * Design, build, and maintain ETL pipelines for ingesting and transforming data into a cloud-based R&D data lake. * Develop reusable data models and libraries to streamline common analytics and ETL use cases. * Implement data quality through validation, monitoring, and anomaly detection mechanisms. * Embed security and compliance through RBAC, data lineage, auditing, and regulatory standards. * Optimize ETL pipelines for performance, scalability, and cost efficiency. * Automate deployments and operations using CI/CD, Infrastructure as Code, and ETL jobs. * Monitor and troubleshoot production pipelines, resolve data and platform issues, and continuously improve reliability. * Master's degree or equivalent practical experience in Data Engineering, Software Engineering, Computer Science, or a related technical field. * Hands-on experience with AWS data lake services, including AWS Glue, S3, Athena, and Lake Formation, covering cataloging, governance, ETL orchestration, and secure data access. * Extensive experience designing and implementing ETL pipelines in Databricks, including migration of cloud-based data lakes and ETL workflows to Databricks. * Strong experience with Delta Lake, including Delta table design, performance optimization, schema evolution, and versioned data management. * Strong experience with cloud-native engineering across AWS/Azure, with a primary focus on AWS data lake environments. * Hands-on experience with Infrastructure as Code, CI/CD, DevOps/MLOps practices, AWS CDK, Ansible, and GitLab CI. * Advanced proficiency in Python and SQL, with the ability to develop robust, maintainable, and reusable code frameworks. * Strong knowledge of observability, data lineage, data quality frameworks, metadata management, and secure data access patterns. * Expertise in cluster and job tuning, job orchestration, storage optimization, and cost management in large-scale data lake environments. * Familiarity with Agile and Scrum methodologies, including sprint planning, backlog refinement, iterative delivery, and cross-functional collaboration. * Strong communication, stakeholder management, problem-solving, and collaboration skills. ## Description * Help design, build and continuously improve the clients online platform. * Research, suggest and implement new technology solutions following best practices/standards. * Take responsibility for the resiliency and availability of different products. * Be a productive member of the team. ## Related Videos - [Building Reliable Serverless Applications with AWS CDK and Testing](https://www.wearedevelopers.com/videos/812-building-reliable-serverless-applications-with-aws-cdk-and-testing) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [The power of Cloud Development Kit (CDK): How to get the most out of it](https://www.wearedevelopers.com/videos/740-the-power-of-cloud-development-kit-cdk-how-to-get-the-most-out-of-it) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Software Developer Salary in The Netherlands [2023]](https://www.wearedevelopers.com/magazine/217-software-developer-salary-in-the-netherlands-2023) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Best Companies in the Netherlands: Top 25 Companies in 2023 ](https://www.wearedevelopers.com/magazine/193-best-companies-in-the-netherlands-top-25-companies-in-2023) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know)