Data Analytical Engineer

Akaasa Technologies
Malvern, PA, United States
about 1 month ago

Role details

Contract type
Temporary to permanent
Employment type
Full-time (> 32 hours)
Compensation
$124,800.0 - $135,200.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Data Analysis Automation of Tests Databases Continuous Integration Extract Transform Load (ETL) Database Development Distributed Computing Environment Github Identity and Access Management
+9 more
Python (Programming Language) Tableau (Software) Test-Driven Development (TDD) Apache Spark Cloudformation Pyspark Data Analytics Functional Programming Cloudwatch

Job description

  • Writes ETL (Extract / Transform / Load) processes, designs database systems, and develops tools for real-time and offline analytic processing.
  • Troubleshoots software and processes for data consistency and integrity. Integrates data from a variety of sources for business partners to generate insight and make decisions.
  • Experience building AWS cloud architecture and supporting services and technologies (Eg: ECS, Glue, S3, Glue Crawler, Redshift, Step Functions, Foundational Services like IAM, Cloud Watch, Cloud formation, Lambda, Secrets Manager, Sagemaker).
  • Translates business specifications into design specifications and code. Responsible for writing programs, ad hoc queries, and reports. Ensures that all code is well structured, includes sufficient documentation, and is easy to maintain and reuse.
  • Partners with internal clients to gain a basic understanding of business functions and informational needs. Gains working knowledge in tools, technologies, and applications/databases in specific business areas and company-wide systems.
  • Participates in all phases of solution development. Explains technical considerations at related meetings, including those with business clients.
  • Expert building Spark data processing applications (Python, Pyspark).
  • Expert with SQL development and Tableau Reporting.
  • Experience with test automation and test-driven development practices.
  • Experience with CI/CD pipeline tools like Github.
  • Tests code thoroughly for accuracy of intended purpose. Reviews end product with the client to ensure adequate understanding. Provides data analysis guidance as required.
  • Provides tool and data support to business users and fellow team members.

Job Responsibilities

  • Troubleshoots software and processes for data consistency and integrity. Integrates data from a variety of sources for business partners to generate insight and make decisions.
  • Experience building AWS cloud architecture and supporting services and technologies (Eg: ECS, Glue, S3, Glue Crawler, Redshift, Step Functions, Foundational Services like IAM, Cloud Watch, Cloud formation, Lambda, Secrets Manager, Sagemaker).

Requirements

  • Malvern, PA Python, PySpark, AWS Strong hands-on experience with PySpark and distributed data processing Experience working with AWS services such as S3, EC2, Lambda, Glue, EMR, Redshift, …

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:21 min

Providing researchers with on-demand analytics environments

Jeremy Murray Jeremy Murray · WWC Europe 2026

3:43 min

The enduring legacy of the amazon S3 storage API

Chris Heilmann +3 · LIVE

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

1:43 min

AWS infrastructure stack and data flow pipeline overview

Artem Volk Artem Volk +1 · WWC 2024

Videos

See all

Related articles

See all