Technical Lead - Data GCP & Databricks

Gapstars
ALMERE, Netherlands
7 days ago
Apply on gapstars.net
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Airflow Data Analysis Automation of Tests BigQuery Code Review Continuous Integration Data Infrastructure Data Transformation Data Flow Control Python (Programming Language) Machine Learning
+17 more
Standard Sql SQL Databases Data Streaming Systems Integration Management of Software Versions Google Cloud SAP Integration Solutions Apache Spark Git Data Layers Data Lakes Data Lineage Machine Learning Operations Terraform Data Pipelines Docker Databricks

Job description

As a Senior Data Engineer, you will work across Google Cloud Platform (GCP) and Databricks, building and evolving scalable data pipelines, data products, and analytics-ready datasets.

The current environment is approximately 60-70% GCP-focused, so strong hands-on GCP experience is mandatory. Initially, your work will primarily focus on the existing GCP platform, with increasing ownership of Databricks pipelines, integrations, and lakehouse capabilities as the platform evolves.

You will also work closely with Data Scientists and Analytics teams to deliver trusted, governed, AI-ready and reporting-ready datasets., 1) GCP Data Engineering

  • Design and maintain scalable data architectures and pipelines on GCP.
  • Build and optimize solutions using BigQuery, Dataflow, Cloud Run, Composer, and GCS.
  • Develop reliable pipelines supporting analytics, reporting, and machine learning workloads.
  • Translate business requirements into scalable technical solutions.
  • Maintain high standards around performance, reliability, security, and cost.

2) Databricks Engineering

  • Own and develop Databricks pipelines and integrations.
  • Work with Delta Lake, Unity Catalog, Workflows, MLflow, and Spark.
  • Support the gradual expansion of Databricks within the wider data platform.
  • Ensure GCP and Databricks workloads follow consistent engineering, governance, and security standards.
  • Help shape future lakehouse architecture and integration patterns.

3) Data Transformation & Semantic Layer

  • Develop transformation workflows using SQL, Python, and dbt.
  • Build and maintain reusable semantic layers and data models.
  • Deliver datasets that are ready for analytics, reporting, AI, and ML use cases.
  • Establish standards for testing, documentation, versioning, and deployment.

4) Data Quality, Governance & Reliability

  • Implement data quality controls and automated testing.
  • Ensure data accuracy, governance, security, and accessibility.
  • Monitor pipeline health, freshness, performance, and operational stability.
  • Troubleshoot incidents and drive continuous improvement.
  • Support data lineage, access controls, and lifecycle management.

5) AI / ML Enablement

  • Work closely with Data Scientists and Analytics teams.
  • Understand ML workflows and the data requirements behind them.
  • Build trusted and reusable datasets for ML and AI use cases.
  • Support ML pipelines and integrations, including the use of MLflow where applicable.
  • Help create AI-ready datasets through the semantic layer.

6) Platform Ownership & Engineering Standards

  • Contribute to infrastructure and deployment standards using Terraform, Docker, CI/CD, and Git.
  • Promote engineering best practices around testing, documentation, code reviews, and operational ownership.
  • Provide technical guidance and mentor other engineers.
  • Document architectures, data flows, integrations, and operational processes.

7) Collaboration & Stakeholder Management

  • Work with Product Owners, Data Scientists, Analysts, Engineers, and business stakeholders.
  • Translate business needs into scalable technical solutions.
  • Communicate technical decisions, trade-offs, risks, and progress clearly.
  • Take ownership of solutions from design through production., Technical Lead - Data GCP & Databricks

Requirements

  • Strong hands-on Google Cloud Platform (GCP) experience is mandatory.
  • Strong experience with BigQuery and GCP-based data pipelines.
  • Nice-to-Have
  • Experience with:
  • Delta Lake
  • Unity Catalog
  • Databricks Workflows
  • MLflow
  • Apache Spark
  • GCP services such as:
  • Dataflow
  • Cloud Run
  • Composer / Airflow
  • Google Cloud Storage
  • Terraform and Docker experience.
  • Experience with SAP integrations.
  • Semantic modeling experience.
  • AI-enabled analytics experience.
  • Retail or e-commerce experience. *

  • Hands-on experience with Databricks in production environments.
  • Strong SQL and Python skills.
  • Experience building and operating modern data platforms or lakehouse environments.
  • Experience with dbt or similar transformation frameworks.
  • Strong understanding of data modeling and semantic layers.
  • Experience supporting ML workflows and Data Science teams.
  • Experience with data quality, governance, security, and monitoring.
  • Experience with CI/CD and Git.
  • Strong ownership and stakeholder communication.

About the company

About Gapstars At Gapstars, we partner with some of Europe’s most ambitious companies, from disruptive tech startups to established Benelux accounting firms, helping them build high-performing teams in engineering and finance.

Headquartered in the Netherlands, with talent hubs in Sri Lanka and Portugal, we are home to 300+ professionals who thrive on solving real-world challenges with modern solutions. Our teams work across domains, from networking, marketplaces, SaaS, and AI to embedded finance and accounting, delivering scalable solutions that drive meaningful outcomes.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on gapstars.net
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all