Senior Data Engineer

Alight, Inc.
United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Airflow Amazon Web Services Amazon S3 Big Data Cloud Computing Cloud Database Cloud Engineering Cloudera Impala Code Review Continuous Integration Extract Transform Load (ETL)
+41 more
Data Profiling Data Warehousing Distributed Computing Environment Distributed Systems Github Apache Hadoop Hadoop Distributed File System Apache Hive Identity and Access Management Subnetting Python (Programming Language) Routing Performance Tuning Shell Script Simple Data Format SQL Databases Sqoop Data Streaming Parquet Data Logging Data Ingestion Apache Yarn Autoscaling Snowflake Apache Spark State Machines Amazon Virtual Private Cloud (VPC) Git Containerization Data Lakes Pyspark Optimization Algorithms AWS Data Analytics Apache Kafka Data Management Interactive Whiteboards Data Pipelines Serverless Computing Docker Programming Languages Control M

Job description

Senior Data Engineer with strong expertise across traditional big-data platforms (Hadoop ecosystem) and modern cloud-native architectures (AWS). Responsible for building scalable, secure, and high-performance data pipelines that span Hadoop clusters and AWS cloud services. Leverages deep knowledge of distributed systems, Spark optimization, cloud automation, and big-data management to support analytics, BI, ML, and AI use cases across the enterprise. Ensures reliability, governance, cost-efficiency, and operational excellence across hybrid data platforms.

Associate should be self-driven, can work with minimal guidance and guide the team technically.

Core Responsibilities

  • Design, build, and maintain high-volume ETL/ELT pipelines across Hadoop (HDFS, Hive, Spark, Kafka) and AWS (Glue, EMR, Lambda, Step Functions, Redshift) .

  • Develop distributed data processing solutions using PySpark, Spark SQL , and scalable cloud serverless patterns.

  • Implement reusable data ingestion frameworks for batch (Sqoop, Hive, Spark) and streaming (Kafka, Kinesis).

  • Optimize data workflows using partitioning, bucketing, compression, file formats (Parquet/ORC).

  • Understanding hybrid data lake architectures using S3 + HDFS , ensuring governance consistency (Atlas, Ranger, Lake Formation).

  • Understanding the reporting requirements and perform data profiling and create design for same.

  • Create data flow diagram and do data modelling.

  • Job orchestration using Airflow, Control-M, Step Functions , or event-driven triggers.

  • Understand auto-scaling, capacity planning, and performance tuning on EMR and Spark clusters.

  • Ensure data is protected and compliant with regulatory standards.

  • Work closely with business stakeholders to enable high-quality datasets.

  • Provide technical leadership in architecture decisions, code reviews, and best-practice adoption and provide technical guidance to peers/juniors in team.

  • Improve reliability, scalability, and performance through automation, autoscaling, and capacity planning.

  • Own deployment, incident response, and post-incident reviews for production environments, troubleshooting Spark performance issues, job failures, and cluster bottlenecks.

  • Understanding security best practices (IAM, KMS, security groups, WAF, parameter/secret management).

  • Optimize cost and usage of AWS resources and recommend architecture improvements.

  • Collaborate closely with developers, QA, and product teams to streamline release processes.

Requirements

Technical Skills

  • Strong experience from 5-8 eyars with the Hadoop ecosystem (HDFS, Hive, Spark, YARN, Kafka).

  • Strong hands-on expertise in Scala, PySpark , Spark optimization techniques, HiveQL, and distributed computing.

  • Good work experience in SQL in hive and impala

  • Good understanding of AWS data stack (S3, Glue, EMR, Lambda, Kinesis, Redshift, Step Functions).

  • Proficiency in at least one scripting/programming language: Python, Shell scripting .

  • Strong experience with CI/CD , GitHub, Git commands.

  • Expertise in ETL and Data Warehousing and cloud concepts.

  • Good understanding of data modelling (star/snowflake), partitioning strategies, and schema evolution.

  • Expertise in data profiling and decision making.

  • Able to understand, design and create data flow diagrams and do data modelling. (knowledge of Miro will be added advantage)

  • Able to understand the architecture and design end-to-end data flow.

  • Hands-on experience with Airflow, Control-M , or other orchestrators.

  • To monitor and support BAU and year end activities, if needed.

  • Well versed with security and compliance aspects in Cloud.

  • Good understanding of AWS networking (VPC, subnets, routing, SGs, NACLs).

  • Familiarity with serverless patterns and containerization (Docker, ECS/EKS).

  • Experience with monitoring/logging tools and incident management practices.

Other Requirements

  • Strong logical and analytical, problem-solving, and communication skills.

  • Communicate effectively and concisely with multiple stakeholders and coordinate and collaborate with cross functional teams.

  • Ability to support both legacy Hadoop workloads and cloud-first architectures.

  • AWS certifications (Data Engineer, Solutions Architect, or Developer) are a plus.

  • Good to have health care domain knowledge.

Benefits & conditions

We offer programs and plans for a healthy mind, body, wallet and life because it’s important our benefits care for the whole person. Options include a variety of health coverage options, wellbeing and support programs, retirement, vacation and sick leave, maternity, paternity & adoption leave, continuing education and training as well as several voluntary benefit options.

By applying for a position with Alight, you understand that, should you be made an offer, it will be contingent on your undergoing and successfully completing a background check consistent with Alight’s employment policies. Background checks may include some or all the following based on the nature of the position: SSN/SIN validation, education verification, employment verification, and criminal check, search against global sanctions and government watch lists, credit check, and/or drug test. You will be notified during the hiring process which checks are required by the position.

Our commitment to Inclusion

We celebrate differences and believe in fostering an environment where everyone feels valued, respected, and supported. We know that diverse teams are stronger, more innovative, and more successful.

At Alight, we welcome and embrace all individuals, regardless of their background, and are dedicated to creating a culture that enables every employee to thrive. Join us in building a brighter, more inclusive future., We offer you a competitive total rewards package, continuing education & training, and tremendous potential with a growing worldwide organization.

About the company

At Alight, we believe a company’s success starts with its people. At our core, we Champion People, help our colleagues Grow with Purpose and true to our name we encourage colleagues to “Be Alight.”

Our Values:

Champion People - be empathetic and help create a place where everyone belongs.

Grow with purpose - Be inspired by our higher calling of improving lives.

Be Alight - act with integrity, be real and empower others.

It’s why we’re so driven to connect passion with purpose. Alight helps clients gain a benefits advantage while building a healthy and financially secure workforce by unifying the benefits ecosystem across health, wealth, wellbeing, absence management and navigation.

With a comprehensive total rewards package, continuing education and training, and tremendous potential with a growing global organization, Alight is the perfect place to put your passion to work.

Join our team if you Champion People, want to Grow with Purpose through acting with integrity and if you embody the meaning of Be Alight.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all