Data Engineer - RHE - Data and Platform Engineer

Roche
Madrid, Spain
11 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English, Spanish
Job source

Tech stack

Agile Methodology Artificial Intelligence Airflow Amazon Web Services Microsoft Azure Big Data Information Engineering Data Infrastructure Extract Transform Load (ETL) Data Stores Data Systems Data Warehousing
+27 more
Database Queries Linux Python (Programming Language) Windows Servers Red Hat Enterprise Linux Power BI Cloud Services Software Engineering SQL Databases Data Streaming Tableau (Software) Talend VMware VSphere Google Cloud Data Storage Technologies System Availability Data Strategy Pandas Pyspark Information Technology Non-relational Database Tools for Reporting Nutanix Data Pipelines Serverless Computing Docker Programming Languages

Job description

We are the core engineering engine for the Data Stores area. Our mission is to accelerate Roche’s Data and AI journey by delivering cutting-edge, scalable database platforms and robust ecosystem and automation capabilities. We don’t just store data; we build the intelligent, self-driving data foundations that power federated data consumption.

As an IT Data Engineer, you will own and optimize core enterprise data systems, driving the architecture and strategy for reliable data pipelines and full-stack infrastructure. You will independently lead initiatives to modernize our platforms, ensure high availability, and deliver scalable, AI-ready data solutions., * Scope / (Content Leadership): Architects and optimizes scalable data pipelines and ETL/ELT frameworks. Leads the maintenance and evolution of complex data infrastructure.

  • Accountability/Problem Solving: Independently identifies, analyzes, and resolves complex data flow and system performance issues, proactively preventing operational bottlenecks.
  • Stakeholder Management: Collaborates with cross-functional stakeholders to define data strategy, ensuring alignment with global data standards and business objectives.
  • Impact/Strategy: Leads key strategic initiatives and high-impact projects, driving technical innovation across the data product line.
  • Complexity / (Product Size) : Engineers and manages robust, high-volume datasets and complex pipelines, ensuring resilience, integrity, and performance of critical data systems.
  • Business / Technical ability: Demonstrates mastery of SQL, Linux, and industrial-grade ETL/ELT tooling, leveraging cloud-native services and AI-driven platforms to scale our data engineering capabilities.

Requirements

  • A bachelor’s degree in Computer Science or equivalent working experience
  • Experience with infrastructure technologies like Windows Server, Red Hat Enterprise Linux, VMware vSphere/Nutanix Acropolis, Backup and Recovery systems, and Relational and non-relational databases
  • Proven experience and mastery in executing complex technical tasks.
  • Knowledge of Agile Frameworks (SAFe) and continuous improvement methodologies would be a plus
  • Proficient in English, and if you speak Spanish, that’s even better!
  • Knowledge of cloud platforms (AWS, Azure, GCP)

Technical Skills

  • Basic knowledge of SQL, Linux, and standard ETL tools, such as Talend, Apache Airflow, and PySpark
  • Growing proficiency in a programming language such as: Pandas, Polars, Prefect, PySpark…; Required: Python
  • Preferred: reporting tools.
  • Familiarity with cloud services, AWS, Azure and GCP, Big Data platforms, and foundational data warehousing concepts.
  • Basic knowledge of Docker containers
  • A reporting tool such as Tableau, Power BI, or ThoughtSpot
  • Software Development Life Cycle best practices

Additional Qualifications

  • Strong collaborative skills, with the ability to work effectively with internal technical teams to ensure continuous data availability.
  • Expertise in optimizing data storage solutions and troubleshooting complex data flow issues independently.
  • AI Code Assistance
  • Demonstrated customer & delivery focus.

About the company

A healthier future drives us to innovate. Together, more than 100’000 employees across the globe are dedicated to advance science, ensuring everyone has access to healthcare today and for generations to come. Our efforts result in more than 26 million people treated with our medicines and over 30 billion tests conducted using our Diagnostics products. We empower each other to explore new possibilities, foster creativity, and keep our ambitions high, so we can deliver life-changing healthcare solutions that make a global impact.

Let’s build a healthier future, together.

Roche is an Equal Opportunity Employer.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · WWC 2024

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:33 min

Refactoring data science workflows using Rapids QDF and Pandas

Paul Graham Paul Graham · LIVE

2:39 min

Experiencing core Linux capabilities for DevOps administration

Michael Cade · LIVE

Videos

See all

Related articles

See all