Data Engineer with AWS & Python

Job Cloud Inc.
McLean, VA, United States
5 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Airflow Amazon Web Services Amazon S3 Cloud Computing Security Cloud Database Databases Continuous Integration Data as a Services Data Validation Information Engineering Data Integration
+36 more
Extract Transform Load (ETL) Data Warehousing Relational Databases Distributed Computing Environment Identity and Access Management Python (Programming Language) Operational Databases Scrum Methodology Cloud Services Standard Sql Data Streaming Unstructured Data Data Logging Data Processing Scripting Data Storage Management Data Ingestion Snowflake Apache Spark Software Troubleshooting AWS Lambda Git Cloudformation Data Lakes Pyspark Kubernetes Information Technology AWS Glue Cloudwatch Restful APIs Terraform Software Version Control Data Pipelines Docker Amazon Redshift Databricks

Job description

We are looking for an experienced Data Engineer with strong AWS and Python expertise to design, develop, and maintain scalable data pipelines and cloud-based data solutions. The ideal candidate will have hands-on experience with AWS data services, Python programming, ETL/ELT development, data integration, and modern data engineering practices., * Design, develop, and maintain scalable data pipelines and ETL/ELT workflows using Python and AWS services.

  • Build and optimize data ingestion and transformation pipelines for structured and unstructured data.
  • Develop reusable Python scripts and applications for data processing, automation, and integration.
  • Work with AWS services such as S3, Glue, Lambda, Redshift, EMR, Athena, Kinesis, and CloudWatch.
  • Implement data processing solutions using PySpark and distributed computing frameworks.
  • Develop data models and optimize data storage and retrieval processes.
  • Perform data quality checks, validation, reconciliation, and error handling.
  • Optimize data pipelines for performance, scalability, reliability, and cost efficiency.
  • Integrate data from APIs, databases, files, and other enterprise data sources.
  • Implement monitoring, logging, and alerting for data pipelines.
  • Collaborate with Data Scientists, Data Analysts, Architects, and business stakeholders.
  • Follow best practices for CI/CD, version control, testing, security, and documentation.
  • Troubleshoot production data issues and provide timely resolution.

Requirements

Required Skills / Must Have

  • 5+ years of experience in Data Engineering or a related field.
  • Strong hands-on experience with Python for data engineering and automation.
  • Strong experience with AWS cloud services, particularly:

  • Amazon S3
  • AWS Glue
  • AWS Lambda
  • Amazon Redshift
  • Amazon Athena
  • Strong knowledge of SQL and relational databases.
  • Experience developing ETL/ELT data pipelines.
  • Experience with PySpark/Spark and distributed data processing.
  • Strong understanding of data warehousing and data lake concepts.
  • Experience with Git and CI/CD pipelines.
  • Strong troubleshooting and analytical skills.

Preferred / Nice to Have

  • Experience with AWS EMR, Kinesis, Step Functions, or MWAA/Airflow.
  • Experience with Terraform or CloudFormation.
  • Knowledge of Databricks, Snowflake, or Delta Lake.
  • Experience working with streaming data pipelines.
  • Knowledge of Docker/Kubernetes.
  • Experience with REST APIs and third-party data integrations.
  • Knowledge of AWS IAM, encryption, and cloud security best practices.
  • Experience with Agile/Scrum methodology.

Education

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field.
  • Equivalent professional experience may be considered.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all