Principal Data Engineer

Roche
Reocín, Spain
10 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
8 years minimum
Working hours
Regular working hours
Languages
Spanish

Tech stack

Airflow Information Engineering Data Governance Data Integration Dataspaces Data Vault Modeling Data Warehousing Digital Assets Dimensional Modeling Distributed Computing Environment Data Flow Control Graph Database
+22 more
Python (Programming Language) MongoDB Neo4j NoSQL Cloud Services SQL Databases Workflow Management Systems Azure Data Factory Apache Spark Data Strategy Cloudformation Containerization Kubernetes Infrastructure Automation Frameworks Apache Flink Cassandra AWS Glue Apache Kafka Terraform Stream Processing Data Pipelines Docker

Job description

Principal Data Engineer We are seeking an exceptionally skilled and influential Principal Data Engineer to join and provide technical leadership for the digital solution navify Cervical Screening.You will be a top?tier expert, driving architectural decisions, overseeing data integration, and ensuring delivery of features and improvements.Key Responsibilities Drive data integration and data fabric adoption across the domains, bringing together data, business and development domains.Ensure best practices and support organization?wide standards for data quality, governance, security, and observability.Act as a technical subject matter expert, mentoring Senior and Staff Engineers and collaborating closely with engineering leadership, data scientists, and business stakeholders.Partner with cross?functional stakeholders to shape data strategy, translate business needs into scalable technical solutions, and influence decision?making at both domain and enterprise levels.Spearhead the evaluation and integration of new technologies to continuously improve the performance and reliability of the data ecosystem.Troubleshoot and resolve the most complex performance and architectural challenges across the data ecosystem.Define and ensure timely delivery of all data?related requirements, including data integration, migration, storage, structure, legal compliance, privacy, quality, and insights generation.Orchestrate DA&R engagement and delivery for digital products, activating the right DA&R capabilities.Act as a key liaison, building and maintaining strong working relationships with cross?functional experts to jointly explore business opportunities and data assets.Required Qualifications 8+ years of professional experience in a Data Engineering role, with significant experience operating at a Principal or Staff?level capacity.Expert proficiency in SQL and Python.Extensive experience architecting and managing large?scale cloud data platforms (e.G., AWS Glue, Google Dataflow, Azure Data Factory) and deep expertise in containerization and orchestration (Docker/Kubernetes).Deep, hands?on expertise in data warehousing concepts, advanced modeling techniques (e.G., dimensional modeling, data vault), and establishing data governance frameworks.Proven track record of building, optimizing, and leading the deployment of massive?scale, highly concurrent data pipelines using distributed processing frameworks (e.G., Apache Spark, Flink).Exceptional communication, negotiation, and influencing skills, with the ability to present complex technical concepts to both technical and executive audiences.Preferred Qualifications Good command of Spanish.Deep experience with workflow orchestration tools (e.G., Apache Airflow, Dagster) and managing them at scale.Experience with NoSQL and graph databases (e.G., MongoDB, Cassandra, Neo4j).Expertise in stream processing architectures and technologies (e.G., Kafka, Kinesis, Pulsar).Extensive experience with Terraform, CloudFormation, or other advanced Infrastructure as Code (IaC) tools and GitOps methodologies.Roche is an Equal Opportunity Employer.#J-*****-Ljbffr

Requirements

Required Qualifications 8+ years of professional experience in a Data Engineering role, with significant experience operating at a Principal or Staff?level capacity. Expert proficiency in SQL and Python. Extensive experience architecting and managing large?scale cloud data platforms (e.G., AWS Glue, Google Dataflow, Azure Data Factory) and deep expertise in containerization and orchestration (Docker/Kubernetes). Deep, hands?on expertise in data warehousing concepts, advanced modeling techniques (e.G., dimensional modeling, data vault), and establishing data governance frameworks. Proven track record of building, optimizing, and leading the deployment of massive?scale, highly concurrent data pipelines using distributed processing frameworks (e.G., Apache Spark, Flink). Exceptional communication, negotiation, and influencing skills, with the ability to present complex technical concepts to both technical and executive audiences. Preferred Qualifications Good command of Spanish. Deep experience with workflow orchestration tools (e.G., Apache Airflow, Dagster) and managing them at scale. Experience with NoSQL and graph databases (e.G., MongoDB, Cassandra, Neo4j). Expertise in stream processing architectures and technologies (e.G., Kafka, Kinesis, Pulsar). Extensive experience with Terraform, CloudFormation, or other advanced Infrastructure as Code (IaC) tools and GitOps methodologies.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:24 min

Comparing Neo4j and GraphQL conceptual models

William Lyon · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all