> Markdown version of [/jobs/ext/3454762-lead-data-engineer](https://www.wearedevelopers.com/jobs/ext/3454762-lead-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Data Engineer - **Company:** EXL SERVICE - **Location:** Jersey City, NJ, United States - **Experience:** Expert - **Salary:** $120,000.0 - $160,000.0 - **Contract:** Permanent contract - **Skills:** Airflow, Amazon Web Services, Amazon S3, Code Review, Information Systems, Information Engineering, Extract Transform Load (ETL), Data Migration, Data Vault Modeling, Data Warehousing, Software Design Patterns, Dimensional Modeling, Distributed Computing Environment, Python (Programming Language), Performance Tuning, Scrum Methodology, Query Optimization, SQL Stored Procedures, SQL Databases, Cloud Platform System, Sql Optimization, Snowflake, Apache Spark, Break Fix, Pyspark, Kubernetes, Information Technology, AWS Glue, Cloudwatch, Data Pipelines - **Published:** September 28, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=27c4389ac1643a10 ## About the Role * Deep experience with Snowflake including data modeling, performance tuning * Proficiency with AWS services - S3, Glue, Lambda, EMR, Redshift, Step Functions, CloudWatch * Strong experience building distributed data processing frameworks with Apache Spark / PySpark * Advanced SQL skills - complex transformations, query optimization, and dimensional modeling * Expertise in DWH design patterns - Kimball, Inmon, Data Vault, star and snowflake schemas * Demonstrated experience leading or contributing to cloud migration and legacy modernization programs * Familiarity with tools such as dbt, Apache Airflow, AWS Glue, or similar orchestration frameworks Solid Python programming for data engineering and automation tasks Qualifications: Experience Requirements * 6-9 years of progressive experience in data engineering * Prior experience in insurance, financial services, or regulated industries preferred * Experience coordinating distributed teams across time zones (onshore/offshore model) * Demonstrated ability to engage with non-technical stakeholders and translate business requirements * Exposure to Agile/Scrum delivery methodology Education * Bachelor's degree in Computer Science, Information Systems, Engineering, or a related field ## Description * Design and implement end-to-end data pipelines using PySpark, Snowflake, and AWS cloud services * Architect scalable ELT/ETL workflows and data warehouse models supporting insurance analytics use cases * Drive data migration and modernization efforts from legacy environments to cloud-native platforms * Develop and review complex SQL transformations, stored procedures, and data quality validation frameworks * Establish and enforce data engineering standards, coding best practices, and pipeline documentation * Provide hands-on troubleshooting and performance optimization across the data stack Team Coordination & Stakeholder Engagement * Coordinate day-to-day activities across onshore and offshore data engineering teams to ensure timely delivery * Serve as a technical point of contact for business stakeholders, translating requirements into engineering deliverables * Facilitate requirement-gathering sessions, sprint planning, and status updates with project teams * Communicate project progress, risks, and dependencies to project managers and client stakeholders * Mentor junior engineers and conduct code reviews to uphold quality standards * Collaborate with data architects, analysts, and QA teams throughout the project lifecycle ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Fully Orchestrating Databricks from Airflow](https://www.wearedevelopers.com/videos/336-fully-orchestrating-databricks-from-airflow) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [How we built an AI-powered code reviewer in 80 hours](https://www.wearedevelopers.com/videos/1511-how-we-built-an-ai-powered-code-reviewer-in-80-hours) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)