Senior Data Engineer

Deloitte T.T.L.
Jersey City, NJ, United States
1 day ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours

Tech stack

Airflow Microsoft Azure Big Data Code Review Information Engineering Extract Transform Load (ETL) Data Systems Apache Hive Python (Programming Language) Microsoft SQL Server Pair Programming Performance Tuning
+15 more
Azure Data Lake Azure Service Bus Azure Data Factory Cloud Monitoring Apache Spark Git Data Lakes Pyspark Information Technology Data Management Terraform Data Pipelines Serverless Computing Key Vault Databricks

Job description

Experteer Overview As a Senior Data Engineer at Deloitte, you will design and implement end-to-end data solutions that enable client decisions at scale. You’ll collaborate with engagement managers and cross-functional teams to deliver robust ETL/ELT pipelines, governance, and performance optimizations in Azure Databricks environments. You’ll mentor engineers, drive design discussions, and evaluate new technologies to modernize clients’ data platforms. This role offers hands-on impact with minimal travel under the Project Delivery Model. Compensation / Benefits * Design, develop and optimize ETL/ELT pipelines using Azure Data Factory and Databricks * Write and tune PySpark / Spark SQL notebooks for large-scale data transformation * Architect end-to-end data solutions across dev UAT prod environments using Unity Catalog * Lead and drive design discussions with client architects and other counterparts * Collaborate with different teams on data contracts and schema agreements * Lead design and optimization of high-volume data pipeline * Define and enforce data engineering standards - naming conventions, partitioning strategies, cluster configurations, Spark tuning * Drive performance optimization - AQE tuning, liquid clustering, broadcast joins, shuffle partition management * Design Databricks cluster policies, autoscaling configurations, and cost optimization strategies * Conduct root cause analysis on production incidents and implement permanent fixes * Mentor junior and mid-level engineers through code reviews and pair programming * Evaluate new technologies and recommend adoption (e.g., DABs, DLT, Auto Loader, Serverless Compute, event hubs) Tasks * Python, PySpark, Spark SQL, SQL Server * Azure (ADF, ADLS Gen2, Key Vault, Azure Monitor) * Databricks (Delta Lake, Unity Catalog, Workflows) * Apache Airflow * Git / Azure DevOps * Deep Spark internals (DAG optimization, spill analysis, skew handling) * Delta Lake advanced features (time travel, deletion vectors, predictive I/O) * Unity Catalog governance (row/column security, external locations, system tables) * IaC - Terraform, Azure ARM templates * Bachelor’s degree in Computer Science or related IT discipline; or equivalent experience Key requirements *

Requirements

and and optimization of high-volume data pipeline * Define and enforce data engineering standards - naming conventions, partitioning strategies, cluster configurations, Spark tuning * Drive performance optimization - AQE tuning, liquid clustering, broadcast joins, shuffle partition management * Design Databricks cluster policies, autoscaling configurations, and cost optimization strategies * Conduct root cause analysis on production incidents and implement permanent fixes * Mentor junior and mid-level engineers through code reviews and pair programming * Evaluate new technologies and recommend adoption (e.g., DABs, DLT, Auto Loader, Serverless Compute, event hubs) Tasks * Python, PySpark, Spark SQL, SQL Server * Azure (ADF, ADLS Gen2, Key Vault, Azure Monitor) * Databricks (Delta Lake, Unity Catalog, Workflows) * Apache Airflow * Git / Azure DevOps * Deep Spark internals (DAG optimization, spill analysis, skew handling) * Delta Lake advanced features (time travel, deletion vectors, a implement I/O) * Unity Catalog governance (row/column security, external locations, system tables) * IaC - Terraform, Azure ARM templates * Bachelor’s degree in Computer Science or related IT discipline; or equivalent experience Key requirements *

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

47 sec

Building modern data pipelines for legacy exports

Dr. Alexander Wachtel Dr. Alexander Wachtel +1 · WWC 2025

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · WWC 2024

Videos

See all

Related articles

See all