Principal Data Engineer

Roche
Reocín, Spain
9 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English, Spanish
Experience level
Senior

Job location

Reocín, Spain

Tech stack

Airflow
Information Engineering
Data Governance
Data Integration
Dataspaces
Data Vault Modeling
Data Warehousing
Digital Assets
Dimensional Modeling
Distributed Computing Environment
Data Flow Control
Graph Database
Python
MongoDB
Neo4j
NoSQL
Cloud Services
SQL Databases
Workflow Management Systems
Azure
Spark
Data Strategy
Cloudformation
Containerization
Kubernetes
Infrastructure Automation Frameworks
Apache Flink
Cassandra
Amazon Web Services (AWS)
Kafka
Terraform
Stream Processing
Data Pipelines
Docker

Job description

Principal Data Engineer We are seeking an exceptionally skilled and influential Principal Data Engineer to join and provide technical leadership for the digital solution navify Cervical Screening.You will be a top?tier expert, driving architectural decisions, overseeing data integration, and ensuring delivery of features and improvements.Key Responsibilities Drive data integration and data fabric adoption across the domains, bringing together data, business and development domains.Ensure best practices and support organization?wide standards for data quality, governance, security, and observability.Act as a technical subject matter expert, mentoring Senior and Staff Engineers and collaborating closely with engineering leadership, data scientists, and business stakeholders.Partner with cross?functional stakeholders to shape data strategy, translate business needs into scalable technical solutions, and influence decision?making at both domain and enterprise levels.Spearhead the evaluation and integration of new technologies to continuously improve the performance and reliability of the data ecosystem.Troubleshoot and resolve the most complex performance and architectural challenges across the data ecosystem.Define and ensure timely delivery of all data?related requirements, including data integration, migration, storage, structure, legal compliance, privacy, quality, and insights generation.Orchestrate DA&R engagement and delivery for digital products, activating the right DA&R capabilities.Act as a key liaison, building and maintaining strong working relationships with cross?functional experts to jointly explore business opportunities and data assets.Required Qualifications 8+ years of professional experience in a Data Engineering role, with significant experience operating at a Principal or Staff?level capacity.Expert proficiency in SQL and Python.Extensive experience architecting and managing large?scale cloud data platforms (e.G., AWS Glue, Google Dataflow, Azure Data Factory) and deep expertise in containerization and orchestration (Docker/Kubernetes).Deep, hands?on expertise in data warehousing concepts, advanced modeling techniques (e.G., dimensional modeling, data vault), and establishing data governance frameworks.Proven track record of building, optimizing, and leading the deployment of massive?scale, highly concurrent data pipelines using distributed processing frameworks (e.G., Apache Spark, Flink).Exceptional communication, negotiation, and influencing skills, with the ability to present complex technical concepts to both technical and executive audiences.Preferred Qualifications Good command of Spanish.Deep experience with workflow orchestration tools (e.G., Apache Airflow, Dagster) and managing them at scale.Experience with NoSQL and graph databases (e.G., MongoDB, Cassandra, Neo4j).Expertise in stream processing architectures and technologies (e.G., Kafka, Kinesis, Pulsar).Extensive experience with Terraform, CloudFormation, or other advanced Infrastructure as Code (IaC) tools and GitOps methodologies.Roche is an Equal Opportunity Employer.#J-*****-Ljbffr

Requirements

Required Qualifications 8+ years of professional experience in a Data Engineering role, with significant experience operating at a Principal or Staff?level capacity. Expert proficiency in SQL and Python. Extensive experience architecting and managing large?scale cloud data platforms (e.G., AWS Glue, Google Dataflow, Azure Data Factory) and deep expertise in containerization and orchestration (Docker/Kubernetes). Deep, hands?on expertise in data warehousing concepts, advanced modeling techniques (e.G., dimensional modeling, data vault), and establishing data governance frameworks. Proven track record of building, optimizing, and leading the deployment of massive?scale, highly concurrent data pipelines using distributed processing frameworks (e.G., Apache Spark, Flink). Exceptional communication, negotiation, and influencing skills, with the ability to present complex technical concepts to both technical and executive audiences. Preferred Qualifications Good command of Spanish. Deep experience with workflow orchestration tools (e.G., Apache Airflow, Dagster) and managing them at scale. Experience with NoSQL and graph databases (e.G., MongoDB, Cassandra, Neo4j). Expertise in stream processing architectures and technologies (e.G., Kafka, Kinesis, Pulsar). Extensive experience with Terraform, CloudFormation, or other advanced Infrastructure as Code (IaC) tools and GitOps methodologies.

Apply for this position