> Markdown version of [/jobs/ext/2294435-data-engineer-aws-spark](https://www.wearedevelopers.com/jobs/ext/2294435-data-engineer-aws-spark). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer (AWS, Spark) - **Company:** Vantor Vault LLC - **Location:** Washington, DC, United States - **Experience:** Expert - **Salary:** $103,000.0 - $140,000.0 - **Contract:** Permanent contract - **Skills:** Microsoft Excel, Artificial Intelligence, Amazon Web Services, Amazon S3, Apache HTTP Server, Information Engineering, Extract Transform Load (ETL), IBM InfoSphere DataStage, Amazon DynamoDB, Python (Programming Language), Key Management, PostgreSQL, Microsoft PowerPoint, SQL Databases, Parquet, Apache Spark, Cloudformation, Pyspark, Information Technology, Amazon Simple Queue Service (SQS), Terraform, Amazon Elastic Mapreduce (EMR) - **Published:** August 29, 2026 - **Apply:** https://www.careerjet.com/jobad/us06dc18ada4d78e13dcc8fbb62a74aafa ## About the Role * 4+ years of data engineering experience * Bachelor's degree * Spark ETL on AWS (Glue, Amazon EMR) in Python and PySpark * S3 data-lake design (Parquet, partitioning, lifecycle) feeding Apache Iceberg tables, Amazon Aurora PostgreSQL, and DynamoDB * Event orchestration (Lambda, Step Functions, SQS/SNS) with secrets management and monitoring * Data quality, validation, and lineage * Infrastructure-as-code (CloudFormation or Terraform) * Basic proficiency in writing, PowerPoint, and Excel Preferred * Master's degree in a relevant field * Trino or comparable federated SQL across the lake and relational stores * Apache Ranger-governed access * Legacy ETL migration (for example DataStage) * Federal information technology or high-volume data experience * Familiarity with AI-assisted developer tooling ## Related Videos - [Parquet, Delta, Iceberg & Ducklake - An introduction for developers](https://www.wearedevelopers.com/videos/100075-parquet-delta-iceberg-ducklake-an-introduction-for-developers) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)