Senior Data Engineer - Castro

SABIA Personal
Crémenes, Spain
1 day ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
4 years minimum
Working hours
Shift work
Languages
English

Tech stack

Query Performance Geographic Information Systems Application Programming Interfaces (APIs) Artificial Intelligence Airflow Amazon Web Services Data Analysis Apache HTTP Server Microsoft Azure Computer Programming Continuous Integration Information Engineering
+24 more
Data Governance Extract Transform Load (ETL) Distributed Computing Environment Interoperability Python (Programming Language) Machine Learning Metadata NetCDF NoSQL Operational Databases Cloud Services Standard Sql Software Engineering SQL Databases User-Centered Design Management of Software Versions Parquet Large Language Models Apache Spark Data Lakes Git Flow Kubernetes Information Technology Data Pipelines

Job description

Our clientis a fast-growing deep-tech company founded in ** and recognized by CB Insights as one of the 100 most promising AI companies globally.They are the largest quantum software company in the EU, with 250+ employees worldwide and growing, delivering advanced solutions trusted by leading global enterprises across several critical industries, including finance, energy, manufacturing, telecom, and industrial sectors.Required Qualifications:· Bachelors or master’s degree in computer science, software engineering, or a related field· 4+ years of professional experience in data engineering, including ownership of production data platforms or pipelines· Expert programming skills in Python and strong command of SQL· Expertise in data modeling, ETL development, and database management, with both SQL and NoSQL databases· Hands-on experience with lakehouse architectures and columnar / open table formats (e.g., Parquet, Apache Iceberg, Delta Lake)· Experience with distributed data processing frameworks such as Spark, and with workflow orchestrators such as Airflow or Argo Workflows· Strong experience with cloud data platforms (Azure, AWS, or GCP), including object storage, containers, and Kubernetes· Solid grounding in data governance: catalogs, metadata, lineage, access control, and dataset versioning· Comfortable with Git-based workflows, CI/CD, and infrastructure-as-code working models· Excellent problem-solving, communication, and collaboration skills; able to lead technical discussions with clients and stakeholders in EnglishNice to have:· Experience with scientific or geospatial data formats and tooling (e.g., Zarr, NetCDF, GRIB2, xarray, H3 spatial indexing)· Experience preparing and serving data for LLM, RAG, or agent-based applications· Previous experience in consulting or client-facing delivery teamsPerks and Benefits:Indefinite contract.Equal pay guaranteed.Variable performance bonus.Signing bonus.They offer work visa sponsorship (If applicable) and relocation package (if applicable).Private health insurance.Eligibility for educational budget according to internal policy.Hybrid opportunity in their offices located in San Sebastian.Flexible working hours.Language classes and discounted lunch options.A high-performance, collaborative environment, operating at pace on cutting-edge technologies.Career plan.Opportunity to learn and teach.Responsibilities· Own the end-to-end design and delivery of data platform architectures - lakehouse, data catalog, and governance - from initial scoping through production release· Design, implement, and operate large-scale ETL/ELT pipelines and workflow orchestration to ensure data is clean, accurate, versioned, and accessible· Define data modeling, partitioning, schema evolution, and versioning conventions so datasets remain queryable, interoperable, and reproducible at scale· Establish and maintain authoritative data catalogs, including schemas, metadata, lineage, sensitivity labels, and access policies· Validate released datasets against their sources for completeness, correctness, schema consistency, and query performance, defining objective acceptance criteria· Work closely with Machine Learning and AI Engineers to make data products directly consumable by analytics, APIs, and AI/agent workflows· Collaborate with clients and cross-functional teams to scope requirements, lead technical sessions, and document architectures for knowledge transfer and internal ownership· Mentor and support other data engineers, reviewing designs and code and raising the team’s engineering standards· Stay up to date with emerging trends in data engineering - open table formats, data catalogs, orchestration - and drive their adoption where they add value

Requirements

· Bachelors or master’s degree in computer science, software engineering, or a related field · 4+ years of professional experience in data engineering, including ownership of production data platforms or pipelines · Expert programming skills in Python and strong command of SQL · Expertise in data modeling, ETL development, and database management, with both SQL and NoSQL databases · Hands-on experience with lakehouse architectures and columnar / open table formats (e.g., Parquet, Apache Iceberg, Delta Lake) · Experience with distributed data processing frameworks such as Spark, and with workflow orchestrators such as Airflow or Argo Workflows · Strong experience with cloud data platforms (Azure, AWS, or GCP), including object storage, containers, and Kubernetes · Solid grounding in data governance: catalogs, metadata, lineage, access control, and dataset versioning · Comfortable with Git-based workflows, CI/CD, and infrastructure-as-code working models · Excellent problem-solving, communication, and collaboration skills; able to lead technical discussions with clients and stakeholders in English Nice to have: · Experience with scientific or geospatial data formats and tooling (e.g., Zarr, NetCDF, GRIB2, xarray, H3 spatial indexing) · Experience preparing and serving data for LLM, RAG, or agent-based applications · Previous experience in consulting or client-facing delivery teams

About the company

Crémenes, León, España

Our clientis a fast-growing deep-tech company founded in ** and recognized by CB Insights as one of the 100 most promising AI companies globally. They are the largest quantum software company in the EU, with 250+ employees worldwide and growing, delivering advanced solutions trusted by leading global enterprises across several critical industries, including finance, energy, manufacturing, telecom, and industrial sectors.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all