Data Engineer (Data Bricks and AWS Cloud)

Powerhouse Institute Inc
United States
3 months ago
Apply on indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
4 years minimum
Compensation
$89,000.0 - $110,000.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Agile Methodology Amazon Web Services Unit Testing Big Data Cloud Computing Cloud Engineering Profiling Databases Continuous Integration Data Validation Data Cleansing
+25 more
Information Engineering Data Integration Extract Transform Load (ETL) Data Transformation Data Security Data Systems Data Warehousing Relational Databases Github Information Sciences Python (Programming Language) Microsoft SQL Server NoSQL Oracle (Applications) Cloud Services Software Engineering Software Repository Data Processing Data Ingestion Information Technology AWS Data Analytics Data Management Data Pipelines Amazon Redshift Databricks

Job description

NOTE: This opportunity is open to W2, 1099 or C2C (no third parties, please). The candidate MUST be authorized to work in the U.S. without sponsorship and able to complete/pass/hold at minimum a public trust investigation. This is a remote opportunity; candidate must be based in the U.S.; have resided in the US for at least 3 years in the past 5 years; ET time zone work schedule.

Daily Responsibilities

  • Play a critical role in ingesting, transforming, and managing data using AWS Cloud Services and Python Programming. Will work directly with senior data engineers o product owners to deliver data products in a collaborative and agile environment.
  • Able to integrate code into a cloud production environment.
  • Work with on-prem database and warehouse solutions such as Oracle or SQL Server or cloud based solutions.
  • Collaborate on data engineering coding principles, standards, designs, frameworks, and chaos testing.
  • Develop data engineering solutions using Python and/or Java.
  • Seek to continuously develop deep AWS engineering skills that optimize code quality and performance.
  • Build data pipelines and other custom automated solutions to speed the ingestion, analysis, and visualization of large volumes of data.
  • Develop data transformation pipelines to clean, enrich, and structure the ingested data for storage in AWS Redshift platform services and storage services.
  • Integrate data from disparate sources to enable seamless data access and analysis within the data management environment.
  • Design and maintain data models within and implement data quality checks and validation processes to ensure data accuracy and integrity.
  • Ensure data security and compliance with government regulations and industry standards.
  • Optimize data pipelines and queries for efficient data processing and retrieval.
  • Work closely with business analysts, data scientists, and other stakeholders to understand data requirements and deliver data solutions that meet business needs.
  • Participate in Agile development sprints including design, development, unit testing and peer review.
  • Prepare comprehensive documentation for data pipelines, data models, and data processes.

Requirements

Do you have experience in Software engineering?, Do you have a Bachelor’s degree?, * Must be authorized to work in the U.S. without sponsorship.

  • Must have resided in the U.S. for at least 3 years within the past 5 years.
  • Must complete/pass/hold at minimum a Public Trust background investigation.
  • 7+ years of progressive professional experience.
  • 5+ years of experience in data engineering, with a focus on data integration, data transformation, and data warehousing.
  • 5+ years cloud development (AWS Cloud) for data processing including experience with ETL processes, including experience with AWS Glue, AWS Data Brew, NoSQL server, SQL Server query experience, AWS Informatica for Relational Database Service (RDS) and the cloud.
  • 4+ years of data bricks experience.
  • Experience in designing and developing Data Pipelines for Data Ingestion or Transformation using AWS technologies.
  • Python programming experience.
  • Experience with data cleansing, profiling, and domain-specific data quality assessment.
  • Experience working in an Agile development environment, with an understanding of Agile principles and practices.
  • Experience working in code repositories (i.e., GitHub) to support CI/CD.
  • Excellent verbal and written communication skills for expressing technical software requirements and designs.
  • Federal and/or healthcare industry experience is a PLUS.
  • Bachelor’s degree in computer science, Information Sciences, or related IT discipline. Additional related professional experience can be substituted for a bachelor’s degree.

Benefits & conditions

Pulled from the full job description

  • AD&D insurance
  • Health insurance
  • 401(k) matching
  • Paid time off
  • Employee discount
  • Vision insurance
  • Dental insurance, Our Comprehensive Employee Benefits Package Includes:
  • 401(k) Retirement Plan (Employer Match)
  • Health Insurance Plans (Medical, Rx, Dental, and Vision - Open Access)
  • Long Term and Short-Term Disability (Company Paid Benefit)
  • Life Insurance (Company Paid Benefit)
  • Employee Assistance Program (EAP)
  • Generous Paid Time Off (PTO)
  • Paid Holidays
  • Voluntary Life and AD&D Insurance
  • Discount Programs for Consumer Products and Wellness Services

Compensation decisions depend on a wide range of factors, including but not limited to skill sets, experience and training, security clearances, licensure and certifications, and other business and organizational needs.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all