Data Engineer - Hadoop to Databricks Migration

Interon IT Solutions LLC
Chantilly, VA, United States
about 2 months ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$107,900.0 - $195,050.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Microsoft Azure Big Data Continuous Integration Data Validation Information Engineering Extract Transform Load (ETL) Data Mapping Data Migration Data Systems Apache Hadoop Hadoop Distributed File System
+14 more
MapReduce Apache Hive Python (Programming Language) Performance Tuning Standard Sql Cloudera Software Deployment SQL Databases Apache Spark Git Data Lakes Pyspark Data Pipelines Databricks

Job description

We are looking for an experienced Senior Data Engineer with strong hands-on experience in Hadoop to Databricks migration. The candidate will be responsible for migrating legacy Hadoop, HDFS, Hive, and Spark workloads to Databricks and modernizing existing data pipelines using PySpark and Delta Lake., * Migrate Hadoop and HDFS workloads to Databricks.

  • Analyze existing Hadoop, Hive, and Spark environments.
  • Migrate large-scale datasets and Hive tables to Databricks.
  • Convert legacy Hive, MapReduce, and Spark jobs to PySpark.
  • Build and maintain ETL/ELT pipelines using Databricks.
  • Develop data solutions using Python, PySpark, Spark SQL, and SQL.
  • Implement Delta Lake and Databricks Lakehouse solutions.
  • Perform source-to-target data mapping and transformation.
  • Validate migrated data for accuracy and completeness.
  • Troubleshoot data migration and pipeline issues.
  • Optimize Spark and PySpark jobs for performance.
  • Support migration testing, production deployment, and post-migration validation.
  • Work with architects and engineering teams during the migration process.
  • Document migration processes and technical solutions.

Requirements

  • 8+ years of Data Engineering or Big Data experience.
  • Strong Hadoop to Databricks migration experience.
  • Strong experience with Hadoop, HDFS, Hive, and Spark.
  • Hands-on experience with Databricks.
  • Strong Python, PySpark, Spark SQL, and SQL skills.
  • Experience converting legacy Hadoop workloads to Databricks.
  • Strong ETL/ELT pipeline development experience.
  • Experience with Delta Lake and Databricks Lakehouse.
  • Experience with data validation and reconciliation.
  • Strong Spark performance tuning experience.
  • Experience with Git and CI/CD.
  • Strong troubleshooting and communication skills.

Preferred Skills

  • Cloudera or Hortonworks to Databricks migration experience.
  • Unity Catalog.
  • Databricks Workflows and Jobs.
  • Auto Loader and Delta Live Tables.
  • AWS, Azure, or GCP experience.
  • Hadoop modernization and decommissioning experience.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all