Principal Data Engineer
Role details
Job location
Tech stack
Job description
Principal Data Engineer We are seeking an exceptionally skilled and influential Principal Data Engineer to join and provide technical leadership for the digital solution navify Cervical Screening.You will be a top?tier expert, driving architectural decisions, overseeing data integration, and ensuring delivery of features and improvements.Key Responsibilities Drive data integration and data fabric adoption across the domains, bringing together data, business and development domains.Ensure best practices and support organization?wide standards for data quality, governance, security, and observability.Act as a technical subject matter expert, mentoring Senior and Staff Engineers and collaborating closely with engineering leadership, data scientists, and business stakeholders.Partner with cross?functional stakeholders to shape data strategy, translate business needs into scalable technical solutions, and influence decision?making at both domain and enterprise levels.Spearhead the evaluation and integration of new technologies to continuously improve the performance and reliability of the data ecosystem.Troubleshoot and resolve the most complex performance and architectural challenges across the data ecosystem.Define and ensure timely delivery of all data?related requirements, including data integration, migration, storage, structure, legal compliance, privacy, quality, and insights generation.Orchestrate DA&R engagement and delivery for digital products, activating the right DA&R capabilities.Act as a key liaison, building and maintaining strong working relationships with cross?functional experts to jointly explore business opportunities and data assets.Required Qualifications 8+ years of professional experience in a Data Engineering role, with significant experience operating at a Principal or Staff?level capacity.Expert proficiency in SQL and Python.Extensive experience architecting and managing large?scale cloud data platforms (e.G., AWS Glue, Google Dataflow, Azure Data Factory) and deep expertise in containerization and orchestration (Docker/Kubernetes).Deep, hands?on expertise in data warehousing concepts, advanced modeling techniques (e.G., dimensional modeling, data vault), and establishing data governance frameworks.Proven track record of building, optimizing, and leading the deployment of massive?scale, highly concurrent data pipelines using distributed processing frameworks (e.G., Apache Spark, Flink).Exceptional communication, negotiation, and influencing skills, with the ability to present complex technical concepts to both technical and executive audiences.Preferred Qualifications Good command of Spanish.Deep experience with workflow orchestration tools (e.G., Apache Airflow, Dagster) and managing them at scale.Experience with NoSQL and graph databases (e.G., MongoDB, Cassandra, Neo4j).Expertise in stream processing architectures and technologies (e.G., Kafka, Kinesis, Pulsar).Extensive experience with Terraform, CloudFormation, or other advanced Infrastructure as Code (IaC) tools and GitOps methodologies.Roche is an Equal Opportunity Employer.#J-*****-Ljbffr
Requirements
Required Qualifications 8+ years of professional experience in a Data Engineering role, with significant experience operating at a Principal or Staff?level capacity. Expert proficiency in SQL and Python. Extensive experience architecting and managing large?scale cloud data platforms (e.G., AWS Glue, Google Dataflow, Azure Data Factory) and deep expertise in containerization and orchestration (Docker/Kubernetes). Deep, hands?on expertise in data warehousing concepts, advanced modeling techniques (e.G., dimensional modeling, data vault), and establishing data governance frameworks. Proven track record of building, optimizing, and leading the deployment of massive?scale, highly concurrent data pipelines using distributed processing frameworks (e.G., Apache Spark, Flink). Exceptional communication, negotiation, and influencing skills, with the ability to present complex technical concepts to both technical and executive audiences. Preferred Qualifications Good command of Spanish. Deep experience with workflow orchestration tools (e.G., Apache Airflow, Dagster) and managing them at scale. Experience with NoSQL and graph databases (e.G., MongoDB, Cassandra, Neo4j). Expertise in stream processing architectures and technologies (e.G., Kafka, Kinesis, Pulsar). Extensive experience with Terraform, CloudFormation, or other advanced Infrastructure as Code (IaC) tools and GitOps methodologies.