Senior Data Engineer - Ods

Banco Santander, S.A.
Madrid, Spain
3 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
5 years minimum
Working hours
Regular working hours
Languages
Spanish

Tech stack

Airflow Amazon Web Services Amazon S3 Data Analysis JIRA Big Data Software Quality Collaborative Software Databases Continuous Integration Information Engineering Extract Transform Load (ETL)
+26 more
Data Stores Data Systems DevOps Distributed Computing Environment Identity and Access Management Python (Programming Language) Machine Learning Cisco Nexus Switches Performance Tuning Cloud Services Standard Sql Scala (Programming Language) SONAR (Symantec) Workflow Management Systems Cloud Platform System Real Time Systems Apache Spark Git Data Lakes Apache Flink Spark Streaming Data Management Stream Processing Splunk Data Pipelines Jenkins

Job description

Experteer Overview Todos los posibles candidatos deben leer con atención los siguientes detalles de este trabajo antes de presentar una candidatura.In this role you design, build and optimize scalable data solutions that power analytics, ML, reporting, and critical decision-making.You will work with large data environments to create robust pipelines and distributed processing using Spark-based technologies.You’ll operate in cloud environments to process, transform and expose data for Analytics, Data Science, and business teams.This position offers a chance to shape data platforms and contribute to a leading digital banking ecosystem.Compensaciones / Beneficios* Design, develop and optimize large-scale data pipelines using Apache Spark (Scala/Spark).* Build batch and near-real-time data processing for high-volume environments.* Create reliable datasets consumed by analytics, data science, ML, and business teams.* Work with cloud data platforms (ideally AWS) to process, transform and store data.* Implement data quality, validation, monitoring and documentation across pipelines.* Collaborate with Data Engineering, Data Science, ML, Architecture and business teams to deliver practical data solutions.* Contribute to engineering best practices around code quality, testing, CI/CD and automation.Responsabilidades* 5+ years in Data Engineering or similar roles.* Experience building production-grade data pipelines in large-scale environments.* Hands-on Spark experience in real projects, preferably Scala.* Experience with cloud data platforms.* Experience in banking, fintech or regulated environments (preferred).* Fluent Spanish; professional English.* Strong Python or Scala coding skills for data engineering tasks.* Solid SQL knowledge and experience with relational/analytic databases.* Experience designing and optimizing ETL/ELT processes.* Experience with AWS or equivalent cloud ecosystems (S3, Glue, Athena, EMR, Redshift, IAM, Lake Formation).* Experience with Git and collaborative software development practices.* Understanding of data quality, validation, monitoring and performance optimization.* Experience with CI/CD or DevOps tools (Jenkins, Sonar, Nexus, Jira, Splunk).* Scala experience (preferred).* Experience with Apache Flink, Spark Streaming or near-real-time applications (preferred).* Experience with Iceberg, Delta Lake or modern lakehouse formats (preferred).xqbhyrx * Experience with Airflow or workflow orchestration tools (preferred).Requisitos principales*

Requirements

  • Contribute to engineering best practices around code quality, testing, CI/CD and automation.Responsabilidades* 5+ years in Data Engineering or similar roles.
  • Experience building production-grade data pipelines in large-scale environments.
  • Hands-on Spark experience in real projects, preferably Scala.
  • Experience with cloud data platforms.
  • Experience in banking, fintech or regulated environments (preferred).
  • Fluent Spanish; professional English.
  • Strong Python or Scala coding skills for data engineering tasks.
  • Solid SQL knowledge and experience with relational/analytic databases.
  • Experience designing and optimizing ETL/ELT processes.
  • Experience with AWS or equivalent cloud ecosystems (S3, Glue, Athena, EMR, Redshift, IAM, Lake Formation).
  • Experience with Git and collaborative software development practices.
  • Understanding of data quality, validation, monitoring and performance optimization.
  • Experience with CI/CD or DevOps tools (Jenkins, Sonar, Nexus, Jira, Splunk).
  • Scala experience (preferred).
  • Experience with Apache Flink, Spark Streaming or near-real-time applications (preferred).
  • Experience with Iceberg, Delta Lake or modern lakehouse formats (preferred). xqbhyrx * Experience with Airflow or workflow orchestration tools (preferred).

About the company

In this role you design, build and optimize scalable data solutions that power analytics, ML, reporting, and critical decision-making. You will work with large data environments to create robust pipelines and distributed processing using Spark-based technologies. You’ll operate in cloud environments to process, transform and expose data for Analytics, Data Science, and business teams. This position offers a chance to shape data platforms and contribute to a leading digital banking ecosystem.Compensaciones / Beneficios* Design, develop and optimize large-scale data pipelines using Apache Spark (Scala/Spark).

  • Build batch and near-real-time data processing for high-volume environments.
  • Create reliable datasets consumed by analytics, data science, ML, and business teams.
  • Work with cloud data platforms (ideally AWS) to process, transform and store data.
  • Implement data quality, validation, monitoring and documentation across pipelines.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

5:47 min

Integrating user stories and test automation via Jira tools

Christoph Ruggenthaler · LIVE

Videos

See all

Related articles

See all