AWS Data Engineer

Code
Newark, NJ, United States
26 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$100,000.0
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Amazon Web Services Amazon S3 Unit Testing Cloud Engineering Code Review Computer Programming Continuous Integration Information Engineering Extract Transform Load (ETL) Data Migration Data Systems
+27 more
Data Warehousing Relational Databases File Systems Amazon DynamoDB Elasticsearch Identity and Access Management Systems Analysis Python (Programming Language) Network Attached Storage (Server Appliance) Standard Sql Shell Script Software Engineering Data Ingestion Apache Spark State Machines AWS Lambda Cloudformation Amazon Relational Database Service Data Lakes Information Technology AWS Data Analytics Apache Kafka SAP Ariba Api Gateway Amazon Simple Queue Service (SQS) Data Pipelines Serverless Computing

Job description

  • Designing, building and maintaining efficient, reusable, and reliable architecture and code.

  • Build reliable and robust Data ingestion pipelines (within AWS, on prem to AWS, etc.)

  • Ensure the best possible performance and quality of high scale data engineering project

  • Participate in the architecture and system design discussions

  • Independently perform hands on development and unit testing of the applications.

  • Collaborate with the development team and build individual components into complex enterprise web systems.

  • Work in a team environment with product, production operation, QE/QA and cross functional teams to deliver a project throughout the whole software development cycle.

  • Responsible to identify and resolve any performance issues

  • Keep up to date with new technology development and implementation

  • Participate in code review to make sure standards and best practices are met.

Pay: Up to $100,000.00 per year

Application Question(s):

  • Are you ready to attend F2F interview (2nd or Final round)?

Requirements

  • Bachelor’s degree in computer science, Software Engineering, MIS or equivalent combination of education and experience

  • Experience implementing, supporting data lakes, data warehouses and data applications on AWS for large enterprises

  • Programming experience with Python, Shell scripting and SQL

  • Solid experience of AWS services such as CloudFormation, S3, Athena, Glue, EMR/Spark, RDS, Redshift, DynamoDB, Lambda, Step Functions, IAM, KMS, SM etc.

  • Solid experience implementing solutions on AWS based data lakes.

  • Should have good experience with AWS Services - API Gateway, Lambda, Step Functions, SQS, DynamoDB, S3, Elasticsearch

  • Serverless application development using AWS Lambda

  • Experience in AWS data lake/data warehouse/business analytics

  • Experience in system analysis, design, development, and implementation of data ingestion pipeline in AWS

  • Knowledge of ETL/ELT

  • End-to-end data solutions (ingest, storage, integration, processing, access) on AWS

  • Architect and implement CI/CD strategy for EDP

  • Implement high velocity streaming solutions using Amazon Kinesis, SQS, and Kafka (preferred)

  • Migrate data from traditional relational database systems, file systems, NAS shares to AWS relational databases such as Amazon RDS, Aurora, and Redshift

  • Migrate data from APIs to AWS data lake (S3) and relational databases such as Amazon RDS, Aurora, and Redshift

  • Implement POCs on any new technology or tools to be implemented on EDP and onboard for real use-case

  • AWS Solutions Architect or AWS Developer Certification preferred

  • Good understanding of Lakehouse/data cloud architecture

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:43 min

The enduring legacy of the amazon S3 storage API

Chris Heilmann Chris Heilmann +3 · LIVE

3:24 min

The governance failures of centralized data lakes

Mario Meir-Huber · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:43 min

AWS infrastructure stack and data flow pipeline overview

Artem Volk Artem Volk +1 · World Congress 2024

2:56 min

Core data lake environment and infrastructure requirements

Christoph Fassbach Christoph Fassbach +1 · World Congress 2024

6:24 min

Distributed data lakes and containerized computing clusters

Ulrich Wurstbauer +1 · LIVE

Videos

See all

Related articles

See all