Databricks Developer

Cognizant
Alcalá de Henares, Spain
12 days ago
Apply on www.adzuna.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Amazon Web Services Amazon S3 Apache HTTP Server Microsoft Azure Big Data Cloud Database Cloud Storage Program Optimization Continuous Integration Data Validation
+29 more
Information Engineering Data Governance Data Infrastructure Extract Transform Load (ETL) Data Vault Modeling Distributed Systems Github Python (Programming Language) Power BI Cloud Services SQL Databases Data Streaming Data Storage Management Multi-Agent Systems Apache Spark Generative AI Backend Git Data Lakes Information Technology Deployment Automation Apache Kafka Data Management Streamlit Framework Terraform Software Version Control Data Pipelines Amazon Redshift Databricks

Requirements

Inscribirse en esta oferta As a Databricks Developer, you will design and build the enterprise data pipelines that power analytics, reporting and AI initiatives for a leading company in the energy sector. Join a fully remote data engineering team working hands-on with cutting-edge Lakehouse technology. HIGH-IMPACT DATA PROJECTS LATEST LAKEHOUSE TECH FULLY REMOTE LEARNING & GROWTH Don’t tick every box? If you meet around 70% of the requirements above, we’d still encourage you to apply. ABOUT THE ROLE We are looking for a highly skilled Data Engineer with 5+ years of experience to design, build and optimize enterprise data pipelines on the Databricks Lakehouse platform for a leading energy sector company. In this role, you will be the hands-on technical driver responsible for transforming raw data into high-quality, actionable datasets. You will build and maintain a Medallion architecture, optimize Spark workloads, and ensure the data infrastructure seamlessly supports advanced analytics, BI dashboards and emerging Generative AI applications. KEY RESPONSIBILITIES Data Pipeline Engineering ? Design, build and maintain scalable, robust ETL/ELT pipelines using Python, SQL and Apache Spark within the Databricks environment. ? Implement and manage a robust Medallion architecture (Bronze, Silver, Gold layers) to process and refine data from diverse sources. ? Develop and maintain the Gold semantic layer specifically optimized for high-performance consumption by BI tools (e.g., Power BI). Platform Optimization & Architecture ? Optimize Databricks workloads, cluster configurations and Spark queries to ensure high performance and cost efficiency. ? Work extensively with open table formats, specifically Delta Lake and Apache Iceberg, to ensure ACID compliance, time travel and efficient data storage. ? Execute complex data migrations, including transitioning legacy workloads from traditional cloud data warehouses (e.g., AWS Redshift) into the Databricks Lakehouse. Data Governance & Automation ? Implement data governance and access control policies at the table, row and column levels using Databricks Unity Catalog. ? Automate deployment processes and pipeline orchestration using Databricks Workflows, CI/CD pipelines (e.g., GitHub Actions, Azure DevOps) and tools like Terraform. ? Embed data quality checks and monitoring directly into pipelines to ensure strict Master Data Management (MDM) standards are upheld. AI & Advanced Analytics Support ? Collaborate closely with Data Scientists and AI Engineers to provision clean, structured data for machine learning model training and inference. ? Support the data foundations required for GenAI frameworks, autonomous agents and AI observability platforms. REQUIRED SKILLS & EXPERIENCE ? 5+ years of dedicated data engineering experience in an enterprise environment. ? Expert-level proficiency in Python and SQL. ? Extensive hands-on experience with Databricks, Apache Spark and Delta Lake. ? Strong understanding of distributed systems, big data architecture and data modeling techniques (e.g., Kimball, Data Vault). ? Deep familiarity with cloud-native data services (AWS, Azure or GCP), specifically cloud storage (S3/ADLS) and compute provisioning. ? Proven experience with version control (Git), CI/CD methodologies and agile software development life cycles. NICE TO HAVE ? Experience evaluating and working with Apache Iceberg alongside Delta Lake. ? Familiarity with streaming data architectures (e.g., Structured Streaming, Kafka). ? Experience building backend frameworks or internal tools using lightweight libraries like Streamlit. EDUCATION ? Bachelor’s or Master’s degree in Computer Science, Information Technology, Data Engineering or a related field. PREFERRED CERTIFICATIONS ? Databricks Certified Data Engineer Associate or Professional. ? AWS, Azure or GCP data/cloud certifications.

About the company

Viseo Madrid, Madrid Volver a la última búsqueda Trabajos ) Databricks Developer ( volver a la última búsqueda Recibir ofertas similares por correo electrónico No gracias, llévame a la oferta de empleo Al crear una alerta, aceptas nuestros Términos y condiciones y Política de privacidad, y el uso de cookies. Inscribirse en esta oferta

Profesiones

  • Técnico
  • Recepciónista
  • Administrador
  • Ventas
  • Enfermero

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

1:34 min

Bringing diverse skills to industrial data science roles

Katja Träumner

Videos

See all

Related articles

See all