AWS Python Data Engineer

Realign Llc
Malvern, PA, United States
11 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$135,000.0
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Airflow Amazon Web Services Amazon S3 Apache HTTP Server Big Data Databases Continuous Integration Data as a Services Data Architecture Data Validation Data Governance
+21 more
Data Integration Extract Transform Load (ETL) Data Systems Data Warehousing DevOps Revision Control Systems Identity and Access Management JSON Python (Programming Language) SQL Databases Parquet Data Processing Cloud Platform System Apache Spark Git Data Lakes Pyspark Apache Kafka Video Streaming Software Version Control Data Pipelines

Job description

  • Design, develop, and maintain scalable ETL/ELT data pipelines.
  • Build and optimize AWS-based data solutions and data architectures.
  • Develop data processing and automation solutions using Python and PySpark.
  • Work with AWS services such as Glue, S3, Lambda, Redshift, EMR, Athena, and Step Functions.
  • Integrate data from APIs, databases, files, and other internal and external sources.
  • Write and optimize SQL queries and support data modeling and data warehousing initiatives.
  • Implement data quality checks and ensure data accuracy, integrity, and reliability.
  • Troubleshoot data pipeline and production issues.
  • Collaborate with data analysts, data scientists, and business stakeholders.
  • Support CI/CD, version control, testing, and deployment processes.

Requirements

We are looking for an experienced AWS Python Data Engineer to design, develop, and support scalable data pipelines and data solutions. The ideal candidate will have strong hands-on experience with Python, SQL, AWS, ETL/ELT processes, and large-scale data processing. Candidates with prior Vanguard experience are highly preferred., * Strong hands-on experience with Python.

  • Strong experience with AWS Data Services.
  • Strong SQL development experience.
  • Experience with ETL/ELT and data pipeline development.
  • Experience with PySpark/Apache Spark.
  • Experience with AWS Glue, S3, Lambda, Redshift, EMR, Athena, and Step Functions.
  • Experience with data lakes, data warehouses, and large-scale data processing.
  • Knowledge of data modeling and data integration.
  • Experience with APIs, JSON, and cloud-based data platforms.
  • Experience with CI/CD and version control tools such as Git.

Preferred Skills

  • Prior Vanguard experience is highly preferred.
  • Experience with Apache Airflow.
  • Experience with Apache Iceberg and Parquet.
  • Knowledge of streaming technologies such as Kafka or Kinesis.
  • Experience with data governance, security, and IAM.
  • Financial services or investment management domain experience.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all