> Markdown version of [/jobs/ext/2900851-software-engineer-data-data-mesh-lakehouse](https://www.wearedevelopers.com/jobs/ext/2900851-software-engineer-data-data-mesh-lakehouse). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer - Data - Data Mesh & Lakehouse - **Company:** Lawrence Harvey - **Location:** Zaragoza, Spain - **Contract:** Permanent contract - **Skills:** Airflow, Apache HTTP Server, Microsoft Azure, Cloud Storage, Code Review, Continuous Integration, Information Engineering, Data Governance, Data Sharing, Distributed Data Store, Machine Learning, Operational Databases, Role-Based Access Control, Software Engineering, SonarQube, Data Streaming, Data Processing, Data Ingestion, Apache Spark, Change Data Capture, Gitlab, Kubernetes, Apache Flink, Data Analytics, Apache Kafka, Data Delivery, Confluent, Databricks - **Published:** September 14, 2026 - **Apply:** https://www.buscojobs.com.es/software-engineer-data-data-mesh-lakehouse-en-zaragoza-ID-371325880 ## About the Role Build and maintain streaming pipelines using technologies such as Kafka, Flink or Confluent.Manage datasets stored across Google Cloud Storage and Azure Blob Storage.Enable secure zero-copy data sharing through technologies such as Databricks Unity Catalog and BigQuery.Implement data governance and access-control models including RBAC and attribute-based access control.Design solutions capable of handling complex schema evolution, including backward and forward compatibility.Tech EnvironmentData & Processing: Apache Spark, Databricks, BigQueryStreaming & Ingestion: Kafka, Flink, Confluent, AirbyteCloud & Storage: GCP, Google Cloud Storage, Azure, Azure Blob StorageOrchestration & Platform: Airflow, KubernetesGovernance: Databricks Unity Catalog, RBAC, ABAC/CBAC, Data ContractsCI/CD: GitLab, Azure DevOps, JFrog ArtifactoryQuality & Security: SonarQube, SnykWhat We're Looking ForStrong professional experience in Data Engineering or Software Engineering focused on data platforms.Hands-on experience with Databricks and Apache Spark.Experience designing distributed data pipelines in cloud environments.Strong understanding of Lakehouse architectures.Experience or strong knowledge of Data Mesh principles and Data Products.Experience with batch and streaming ingestion patterns.Knowledge of CDC and incremental data processing strategies.Experience dealing with schema evolution in production data pipelines.Understanding of modern data governance and access-control models.Experience with Airflow, Kubernetes and CI/CD. xsgfvud Comfortable working collaboratively through code reviews and technical design discussions. ## Description Software Engineer - Data | Data Mesh & Lakehouse¿Todo listo para enviar su solicitud?Por favor, lea la descripción al menos una vez antes de hacer clic en "Solicitar".About the RoleWe are looking for a Software Engineer - Data to join a central Data Delivery team building the foundation of a global Data Mesh platform based on a Lakehouse architecture.You will work on the engineering layer responsible for ingesting, processing and provisioning source-aligned data products into central data catalogs, enabling teams across the organisation to consume trusted data for analytics, reporting, machine learning and other data-driven applications.A key part of the platform is enabling scalable data consumption through zero-copy data sharing, while maintaining strong governance, security and data quality standards.What You'll Be Working OnDesign and build scalable batch and streaming data pipelines.Develop ingestion solutions for source-aligned data products.Work with Apache Spark and Databricks within a modern Lakehouse environment.Implement data ingestion strategies including Full Loads, Delta Loads and Change Data Capture (CDC).Build and maintain streaming pipelines using technologies such as Kafka, Flink or Confluent.Manage datasets stored across Google Cloud Storage and Azure Blob Storage.Enable secure zero-copy data sharing through technologies such as Databricks Unity Catalog and BigQuery.Implement data governance and access-control models including RBAC and attribute-based access control.Design solutions capable of handling complex schema evolution, including backward and forward compatibility.Tech EnvironmentData & Processing: Apache Spark, Databricks, BigQueryStreaming & Ingestion: Kafka, Flink, Confluent, AirbyteCloud & Storage: GCP, Google Cloud Storage, Azure, Azure Blob StorageOrchestration & Platform: Airflow, KubernetesGovernance: Databricks Unity Catalog, RBAC, ABAC/CBAC, Data ContractsCI/CD: GitLab, Azure DevOps, JFrog ArtifactoryQuality & Security: SonarQube, SnykWhat We're Looking ForStrong professional experience in Data Engineering or Software Engineering focused on data platforms.Hands-on experience with Databricks and Apache Spark.Experience designing distributed data pipelines in cloud environments.Strong understanding of Lakehouse architectures.Experience or strong knowledge of Data Mesh principles and Data Products.Experience with batch and streaming ingestion patterns.Knowledge of CDC and incremental data processing strategies.Experience dealing with schema evolution in production data pipelines.Understanding of modern data governance and access-control models.Experience with Airflow, Kubernetes and CI/CD.xsgfvud Comfortable working collaboratively through code reviews and technical design discussions. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Fully Orchestrating Databricks from Airflow](https://www.wearedevelopers.com/videos/336-fully-orchestrating-databricks-from-airflow) - [Let's Get Started With Apache Kafka® for Python Developers](https://www.wearedevelopers.com/videos/565-let-s-get-started-with-apache-kafka-for-python-developers) - [GitLab CI pipelines for a whole company](https://www.wearedevelopers.com/videos/143-gitlab-ci-pipelines-for-a-whole-company) - [Inside Bitpanda's Tech Stack: Scaling a European Fintech Leader - Markus Dorner](https://www.wearedevelopers.com/videos/1979-inside-bitpanda-s-tech-stack-scaling-a-european-fintech-leader-markus-dorner) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)