Spark Technical Lead

Cloudera
Valencia, Spain
3 days ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
6 years minimum
Working hours
Regular working hours

Tech stack

Java (Programming Language) Apache HTTP Server Information Engineering Data Infrastructure Software Debugging Distributed Computing Environment Distributed Systems Fault Tolerance Python (Programming Language) Open Source Technology Cloudera Scala (Programming Language)
+6 more
SQL Databases System Testing Parquet Data Processing Apache Spark Build Management

Job description

At Cloudera, we empower people to transform complex data into clear and actionable insights. With as much data under management as the hyperscalers, we’re the preferred data partner for the top companies in almost every industry. Powered by the relentless innovation of the open source community, Cloudera advances digital transformation for the world’s largest enterprisesThe Data Platform Pillar is the bedrock of Cloudera’s technology, where we design and build the core components that let our customers store, manage, and process data with unmatched scalability, security, and performanceCloudera is seeking a Senior Staff Software Engineer, Spark (Java) with strong distributed systems expertise to work on the Cloudera distribution of Apache Spark and Livy. The role involves building enterprise-grade systems for customers running Spark on thousands of nodes and processing petabytes of dataWe are looking for a passionate engineer eager to enhance a product already supporting major production systems and to drive the next-generation Data Engineering experience. You will collaborate with a distributed team across the United States and Hungary, including multiple Apache Spark committersDesign new features for Cloudera’s data engineering experience, and take them from prototypes to leading a team to deliver the feature in production at scaleContribute to Apache Spark, LivyDevelop new features in Scala/Java/Python on a modern platformsGain expertise in distributed data processing, from SQL planners and optimizers, to data layout and table formats like Apache Parquet and Iceberg, to fault tolerance in distributed systemsGain a solid understanding and deep technical knowledge of components across the Cloudera Data Engineering Experience stack, but focusing on Iceberg and Spark, which you can utilize in your daily tasksGet to work on large scale distributed systems, from 100s to **s of nodes, in production clustersDebug system level deployment issues, root cause analysis, perform system test analysis and resolve failuresWork on improving internal infrastructureCollaborate with other team members and stakeholdersBenefitsComprehensive medical, dental, and vision plansOn-site gyms at various locationsGym reimbursementPaid medical disability & child bonding leavesVolunteer time off policyQuarterly team building offsitesTuition reimbursementFree and catered lunches based on office locationDiscount programsBusiness travel insuranceWork from home opportunities6+ years professional software developmentExperience with distributed systemsExperience with systems design, developmentBsc/Msc in related field or equivalent experienceStrong oral and written communication skillsPassionate about programming, clean coding habits, attention to detail, and focus on quality(Most importantly) Open-minded, desire to learn new things and build great productsWe use Java/Scala/Python in projects, you should have a strong understanding of at least one of the following languages: Java, Scala, Python. And interested to learn the languages we’re usingStrong ability to research and solve problems independently without constant supervisionExperience leading and delivering complex product enhancementsExperience with SQL plannersExperience with using/developing Apache Spark, Livy or other related technologiesExperience with large-scale, distributed systems design and development with an understanding of scaling, performance, and schedulingSolid experience with at least one cloud s#J-***-Ljbffr

Requirements

6+ years professional software developmentExperience with distributed systemsExperience with systems design, developmentBsc/Msc in related field or equivalent experienceStrong oral and written communication skillsPassionate about programming, clean coding habits, attention to detail, and focus on quality(Most importantly) Open-minded, desire to learn new things and build great productsWe use Java/Scala/Python in projects, you should have a strong understanding of at least one of the following languages: Java, Scala, Python. And interested to learn the languages we’re usingStrong ability to research and solve problems independently without constant supervisionExperience leading and delivering complex product enhancementsExperience with SQL plannersExperience with using/developing Apache Spark, Livy or other related technologiesExperience with large-scale, distributed systems design and development with an understanding of scaling, performance, and schedulingSolid experience with at least one cloud s #J-*****-Ljbffr

Benefits & conditions

Comprehensive medical, dental, and vision plans On-site gyms at various locations Gym reimbursement Paid medical disability & child bonding leaves Volunteer time off policy Quarterly team building offsites Tuition reimbursement Free and catered lunches based on office location Discount programs Business travel insurance

About the company

At Cloudera, we empower people to transform complex data into clear and actionable insights. With as much data under management as the hyperscalers, we’re the preferred data partner for the top companies in almost every industry. Powered by the relentless innovation of the open source community, Cloudera advances digital transformation for the world’s largest enterprises

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

2:50 min

How Parquet metadata enables efficient data reading

Matthias Niehoff Matthias Niehoff · World Congress 2026 Europe

1:41 min

Visualizing the complex developer journey for JVM ecosystems

Bobur Umurzokov · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

2:04 min

Comparing offline data analytics with online stream processing

Artem Volk Artem Volk +1 · World Congress 2024

Videos

See all

Related articles

See all