Big Data / AWS Data Engineer

VDart, Inc.
McLean, VA, United States
9 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Microsoft Excel Airflow Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Apache HTTP Server Microsoft Azure Big Data Code Review Customer Data Management Data Architecture Information Engineering
+37 more
Data Governance Extract Transform Load (ETL) Data Transformation Data Sharing Data Warehousing Dimensional Modeling Distributed Computing Environment Apache Hive JSON Python (Programming Language) MicroStrategy MySQL Performance Tuning Query Optimization Cloud Services Shell Script SQL Stored Procedures PL-SQL SQL Databases Extensible Markup Language (XML) Parquet Data Processing Snowflake Apache Spark AWS Lambda Pyspark Semi-structured Data Avro AWS Glue AWS Data Analytics Apache Kafka Data Management Functional Programming Data Pipelines Serverless Computing Amazon Redshift Databricks

Job description

  • Design, develop, and maintain scalable data pipelines using Spark (Scala/Python), AWS Glue, AWS Lambda, and Shell scripting.
  • Build and support ETL/ELT processes across AWS platforms, including S3, Hive, Apache Iceberg, and Amazon Redshift.
  • Develop and maintain Redshift SQL/PLSQL code, stored procedures, and data-processing frameworks.
  • Perform query tuning, performance optimization, and data-sharing implementations.
  • Support ingestion and processing of structured and semi-structured data formats including JSON, CSV, XML, ORC, Parquet, and Avro.
  • Design, enhance, and support solutions that provide a unified view of customer data across systems.
  • Integrate multiple data sources, applications, and business processes to support reporting, analytics, and customer engagement initiatives.
  • Collaborate with business and technical teams to define and implement scalable data solutions.
  • Manage and support Data Subject Requests (DSRs), including customer data erasure requests.
  • Work closely with the DPO team to ensure compliance with privacy regulations and internal policies.
  • Maintain and support the MicroStrategy Privacy Compliance module. (Very infrequent once in year or 2 years)
  • Contribute to data governance initiatives that improve data quality, consistency, and compliance.
  • Analyze and troubleshoot complex data issues using Redshift SQL, MySQL, Python/PySpark, Shell scripting, and Excel.
  • Investigate data mismatches, reporting issues, and business data anomalies.
  • Support business analysis related to customer tiers, loyalty programs, and email marketability.
  • Provide production support and root-cause analysis for critical data platform issues.
  • Technical Leadership & Collaboration
  • Participate in Customer 360, Privacy, Data Governance, and Phaedon initiatives.
  • Conduct code reviews and design reviews.
  • Provide cross-team technical support and problem resolution.
  • Research, design, and develop next-generation data capabilities and products.
  • Act as a subject matter expert for customer data platforms.
  • Primary Technologies
  • AWS: S3, EC2, Redshift (RA3/Serverless), S3 Tables, S3 Hive, AWS Glue, Lambda (Python), Kafka, Airflow, Apache Iceberg
  • Spark (Scala/Python)
  • Redshift SQL/PLSQL, MySQL
  • Data Modeling and Query Optimization
  • Key Strengths
  • Strong ownership and accountability
  • Strong AWS Data Engineering stack awareness
  • Ability to bridge business and technical teams
  • Expertise across engineering, analytics, compliance, architecture, and operational support
  • Adaptability and willingness to take on diverse data-related challenges, Lead Data Engineer (Python, AWS, Spark, Kafka, SQL, Snowflake, Databricks, GenAI) Do you love building and pioneering in the technology space? Do you enjoy solving complex busine…
  • 16 hours ago, Lead Data Engineer (Python, AWS, Spark, Kafka, SQL, Snowflake, Databricks, GenAI) Do you love building and pioneering in the technology space? Do you enjoy solving complex busine…
  • 19 hours ago, Lead Data Engineer (Python, AWS, Spark, Kafka, SQL, Snowflake, Databricks, GenAI) Do you love building and pioneering in the technology space? Do you enjoy solving complex busine…
  • 19 hours ago +

Requirements

  • 12+ years of experience in Data Warehouse / Big Data ecosystems.
  • 5+ years in a lead /SME role

Strong understanding of:

  • Data architecture & dimensional modeling
  • ETL/ELT frameworks
  • Distributed data processing
  • Cloud data platforms
  • Data governance and security
  • Experience managing large-scale enterprise data environments.
  • Proven track record of delivering complex data programs.
  • Excellent communication and stakeholder management abilities., * Experience in large enterprise or global delivery environments.
  • Exposure to data modernization and cloud transformation programs.
  • Certifications in cloud platforms (AWS/Azure/GCP) or data engineering.
  • Experience in regulated industries (Hospitality, Finance, Healthcare, Telecom, etc.).

Key Competencies:

  • Strategic thinking with hands-on technical depth
  • Executive presence and client-facing confidence
  • Problem-solving and decision-making ability
  • Operational excellence mindset

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

3:02 min

Audience Q&A on data formats and engine tradeoffs

Matthias Niehoff Matthias Niehoff · WWC Europe 2026

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

1:48 min

Analyzing network packets with database protocol tools

Daniël van Eeden Daniël van Eeden · WWC Europe 2026

1:52 min

Customizing block storage tiers and formats

Ricardo Sueiras Sueiras · LIVE

Videos

See all

Related articles

See all