Data Engineer - Hadoop to Databricks Migration
Interon IT Solutions LLC
Chantilly, VA, United States
about 2 months ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.careerjet.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$107,900.0 - $195,050.0
Working hours
Regular working hours
Job source
Tech stack
Amazon Web Services
Microsoft Azure
Big Data
Continuous Integration
Data Validation
Information Engineering
Extract Transform Load (ETL)
Data Mapping
Data Migration
Data Systems
Apache Hadoop
Hadoop Distributed File System
+14 more
MapReduce
Apache Hive
Python (Programming Language)
Performance Tuning
Standard Sql
Cloudera
Software Deployment
SQL Databases
Apache Spark
Git
Data Lakes
Pyspark
Data Pipelines
Databricks
Job description
We are looking for an experienced Senior Data Engineer with strong hands-on experience in Hadoop to Databricks migration. The candidate will be responsible for migrating legacy Hadoop, HDFS, Hive, and Spark workloads to Databricks and modernizing existing data pipelines using PySpark and Delta Lake., * Migrate Hadoop and HDFS workloads to Databricks.
- Analyze existing Hadoop, Hive, and Spark environments.
- Migrate large-scale datasets and Hive tables to Databricks.
- Convert legacy Hive, MapReduce, and Spark jobs to PySpark.
- Build and maintain ETL/ELT pipelines using Databricks.
- Develop data solutions using Python, PySpark, Spark SQL, and SQL.
- Implement Delta Lake and Databricks Lakehouse solutions.
- Perform source-to-target data mapping and transformation.
- Validate migrated data for accuracy and completeness.
- Troubleshoot data migration and pipeline issues.
- Optimize Spark and PySpark jobs for performance.
- Support migration testing, production deployment, and post-migration validation.
- Work with architects and engineering teams during the migration process.
- Document migration processes and technical solutions.
Requirements
- 8+ years of Data Engineering or Big Data experience.
- Strong Hadoop to Databricks migration experience.
- Strong experience with Hadoop, HDFS, Hive, and Spark.
- Hands-on experience with Databricks.
- Strong Python, PySpark, Spark SQL, and SQL skills.
- Experience converting legacy Hadoop workloads to Databricks.
- Strong ETL/ELT pipeline development experience.
- Experience with Delta Lake and Databricks Lakehouse.
- Experience with data validation and reconciliation.
- Strong Spark performance tuning experience.
- Experience with Git and CI/CD.
- Strong troubleshooting and communication skills.
Preferred Skills
- Cloudera or Hortonworks to Databricks migration experience.
- Unity Catalog.
- Databricks Workflows and Jobs.
- Auto Loader and Delta Live Tables.
- AWS, Azure, or GCP experience.
- Hadoop modernization and decommissioning experience.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.careerjet.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
DS
Dhannush Subramani
about 4 years ago
BB
Benedikt Bischof
Making Data Warehouses Fast: A Developer’s Story
about 4 years ago
EM
Eli McGarvie
Data Engineer Salary UK
about 3 years ago
CH
Chris Heilmann
Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production
almost 2 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
LM
Luis Minvielle
How to Become an AI Engineer
almost 3 years ago