Sr. Data Engineer

Akaasa Technologies
Chicago, IL, United States
1 day ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Temporary to permanent
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Airflow Amazon Web Services Amazon S3 Apache HTTP Server Information Systems Computer Programming Continuous Integration Data as a Services Extract Transform Load (ETL) Data Mining Data Warehousing
+17 more
Distributed Computing Environment Github Python (Programming Language) Scrum Methodology Cloud Services Data Streaming Apache Spark Reliability of Systems Cloudformation Data Lakes Pyspark Information Technology AWS Glue Apache Kafka Terraform Data Pipelines Databricks

Requirements

Technical Stack & Requirements: Core Competencies: Python, Apache Spark (PySpark), AWS Glue ETL, Data Lake concepts (Medallion Architecture), Databricks, Sagemaker. Infrastructure: CloudFormation/Terraform, CI/CD with GitHub Actions. Core Responsibilities: Orchestrate data extraction from diverse legacy and modern sources, engineer high-performance scalable pipelines, and provide expert-level production support to ensure system reliability. Qualifications: Bachelor’s degree in Computer Science or related field. Strong hands-on experience with Apache Spark and Glue. Proven ability to work with Python-based data pipelines. Experience with Git and GitHub Actions is mandatory. Required Qualifications Bachelor’s degree in Computer Science, Engineering, Information Systems, or related discipline. 5-7+ years of enterprise Data Engineering experience. Strong Python development experience. Advanced SQL programming skills. Extensive experience with Apache Spark and PySpark. Strong experience with Databricks application. Hands-on experience with AWS data services including: o S3 o AWS Glue o Athena o Lambda o Step Functions o EventBridge Experience building enterprise ETL and ELT pipelines. Strong understanding of distributed data processing. Experience working with large-scale cloud data platforms. Strong knowledge of dimensional data modeling. Experience implementing CI/CD pipelines. Familiarity with Agile/Scrum methodologies. Excellent communication and stakeholder management skills. Preferred Qualifications Financial Services or Asset Management industry experience. Knowledge of Apache Iceberg, Delta Lake, or Hudi. Experience with Airflow. Knowledge of Kafka or Kinesis streaming. Experience supporting AI/ML and advanced analytics platforms. AWS Professional or Specialty Certifications.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all