Lead Data Engineer- AWS

Ecloud Labs
United States
14 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Artificial Intelligence Airflow Amazon Web Services Data Analysis Big Data Computer Programming Continuous Integration Data Systems Python (Programming Language) Data Streaming Large Language Models
+7 more
Apache Spark Containerization Data Lakes Kubernetes Infrastructure Automation Frameworks Apache Flink Apache Kafka

Job description

As part of the Big Data Engineering and Analytics organization, this role will focus on enabling a modern data lake and data mesh architecture that supports high volume, high velocity, and high variety data across enterprise and platform domains.

The Lead Data Engineer will bring deep hands-on expertise in Kubernetes, Kafka streaming platforms. This role partners closely with data architects, analytics teams, and platform engineering to deliver reliable, governed, and high-quality data solutions.

Responsibilities

  • Lead the design and implementation of scalable big data pipelines for batch and real time processing

  • Build and operate streaming data platforms using Kafka

  • Design and deploy cloud native data solutions across AWS

  • Develop and manage containerized workloads using Kubernetes

  • Enable data mesh architecture with domain oriented data products

  • Design and implement data lake and lakehouse architectures

  • Ensure data quality, reliability, and observability

  • Implement governance capabilities including metadata and lineage

  • Collaborate with analytics, AI, and business teams

  • Optimize performance and cost efficiency

Requirements

10+ years of experience in big data engineering

  • Strong experience with Kafka

  • Strong experience with Kubernetes

  • Experience with AWS

  • Experience building data lakes or lakehouse platforms

  • Experience with data mesh concepts

  • Strong programming skills in Python, Scala, or Java

  • Experience with Spark, Flink, or Beam

  • Experience with Airflow or orchestration tools

  • Understanding of data modeling and governance

  • Experience with CICD and infrastructure automation

  • Experience supporting AI and machine learning workloads including large language models such as Claude or similar platforms

  • Strong communication and leadership skills

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:24 min

The governance failures of centralized data lakes

Mario Meir-Huber · LIVE

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

6:24 min

Distributed data lakes and containerized computing clusters

Ulrich Wurstbauer +1 · LIVE

1:43 min

AWS infrastructure stack and data flow pipeline overview

Artem Volk Artem Volk +1 · WWC 2024

Videos

See all

Related articles

See all