> Markdown version of [/jobs/ext/1985230-lead-aws-data-engineer](https://www.wearedevelopers.com/jobs/ext/1985230-lead-aws-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead AWS Data Engineer - **Company:** Centraprise LLC - **Location:** Jersey City, NJ, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Airflow, Amazon Web Services, Amazon S3, Code Review, Information Engineering, Extract Transform Load (ETL), Data Migration, Data Vault Modeling, Data Warehousing, Software Design Patterns, Dimensional Modeling, Distributed Computing Environment, Python (Programming Language), Performance Tuning, Scrum Methodology, Query Optimization, SQL Stored Procedures, SQL Databases, Cloud Platform System, Sql Optimization, Snowflake, Apache Spark, Break Fix, Pyspark, Kubernetes, AWS Glue, AWS Data Analytics, Cloudwatch, Data Pipelines - **Published:** August 8, 2026 - **Apply:** https://www.careerjet.com/jobad/us47deeaffef5463bdc41010d3d1e2c6d2 ## About the Role * Deep experience with Snowflake including data modeling, performance tuning * Proficiency with AWS services - S3, Glue, Lambda, EMR, Redshift, Step Functions, CloudWatch * Strong experience building distributed data processing frameworks with Apache Spark / PySpark * Advanced SQL skills - complex transformations, query optimization, and dimensional modeling * Expertise in DWH design patterns - Kimball, Inmon, Data Vault, star and snowflake schemas * Demonstrated experience leading or contributing to cloud migration and legacy modernization programs * Familiarity with tools such as dbt, Apache Airflow, AWS Glue, or similar orchestration frameworks * Solid Python programming for data engineering and automation tasks, * 6 9 years of progressive experience in data engineering * Prior experience in insurance, financial services, or regulated industries preferred * Experience coordinating distributed teams across time zones (onshore/offshore model) * Demonstrated ability to engage with non-technical stakeholders and translate business requirements Exposure to Agile/Scrum delivery methodology ## Description We are seeking an experienced Lead Data Engineer to support complex data engineering initiatives within our insurance data and analytics practice. This role combines Deep technical expertise with strong coordination skills, working closely with onshore and offshore teams, business stakeholders, and project leadership to deliver enterprise data modernization and migration programs. The candidate will serve as a technical point of contact for cross-functional teams while remaining hands-on with cloud data technologies., Technical Delivery * Design and implement end-to-end data pipelines using PySpark, Snowflake, and AWS cloud services * Architect scalable ELT/ETL workflows and data warehouse models supporting insurance analytics use cases * Drive data migration and modernization efforts from legacy environments to cloud-native platforms * Develop and review complex SQL transformations, stored procedures, and data quality validation frameworks * Establish and enforce data engineering standards, coding best practices, and pipeline documentation * Provide hands-on troubleshooting and performance optimization across the data stack Team Coordination & Stakeholder Engagement * Coordinate day-to-day activities across onshore and offshore data engineering teams to ensure timely delivery * Serve as a technical point of contact for business stakeholders, translating requirements into engineering deliverables * Facilitate requirement-gathering sessions, sprint planning, and status updates with project teams * Communicate project progress, risks, and dependencies to project managers and client stakeholders * Mentor junior engineers and conduct code reviews to uphold quality standards * Collaborate with data architects, analysts, and QA teams throughout the project lifecycle ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [How we built an AI-powered code reviewer in 80 hours](https://www.wearedevelopers.com/videos/1511-how-we-built-an-ai-powered-code-reviewer-in-80-hours) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Dev Digest 162: AI careers, MCP, AWS best practices & floppy sweaters](https://www.wearedevelopers.com/magazine/571-dev-digest-162-ai-careers-mcp-aws-best-practices-floppy-sweaters) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path)