Senior Data Engineer

Akumin Inc.
United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$125,000.0 - $150,000.0
Working hours
Regular working hours
Job source

Tech stack

JavaScript (Programming Language) Agile Methodology Artificial Intelligence Amazon Web Services Amazon S3 Microsoft Azure Big Data BigQuery Clinical Data Repository Data Files Data Governance Data Integration
+25 more
Extract Transform Load (ETL) Data Mining Data Stores Data Warehousing Dicom Apache Hadoop Python (Programming Language) Reference Data Cloud Services SQL Databases Transact-SQL Enterprise Data Management Data Processing Scripting Freeform SQL Data Storage Technologies Fast Healthcare Interoperability Resources Snowflake Apache Spark Electronic Medical Records Containerization Information Technology Data Management Data Pipelines Docker

Job description

The Senior Data Engineer will play a crucial role in building out the company’s enterprise data platform to support analytics and AI. As part of the Enterprise Data team, you will be tasked with the responsibility of developing quality data collection processes, maintaining the integrity of our data foundations and enabling business leaders and data scientists across the company to have rapid access to the data they need for decision-making and innovation. The role will be responsible for the infrastructure and operations of the organizations data storage solutions, third-party integrations data provisioning services and overall ETL/ELT code-based solutions and tooling that process data.

Specific duties include, but are not limited to:

  • Development of complex SQL queries for data extraction, manipulation, and reporting and transformation technologies

  • Design and implement robust ETL/ELT pipelines using custom-tooling (Python/Google) and off-the shelf tooling with focus on monitoring, supportability, and resource stewardship.

  • Contribute to and leverage coding standards and best practices to ensure efficient and re-usable services and components.

  • Architect, implement and deploy new data models and data processes in production

  • Builds data pipelines which acquire, cleanse, transform and publish data from a wide variety of sources.

  • Assembles large, complex data sets which meets functional and non-functional business requirements.

  • Partners with data asset managers, architects, and development leads to ensure technical solution provides data which is fit for use and in line with architecture blueprints.

  • Identify, document, design, and implement internal process improvements.

  • Other duties as assigned by management.

Requirements

  • Bachelor’s degree required in Data Science, Computer Science or MIS, Mathematics, Engineering, or related field.

  • 5+ years of prior experience in Data Management / ETL / ELT / Data Warehousing.

  • Experience in writing Data Quality routines for cleansing of data and capturing confidence score and master data management (MDM).

  • Hands-on experience with designing and implementing data pipelines and ELT/ETL processes (ex. Fivetran, DBT).

  • Hands-on experience with cloud platforms (ex.GCP, AWS, Azure) and related services (ex. BigQuery, S3, Snowflake, etc.).

  • Strong understanding of data modeling, data integration, and data governance principles (ex. DBT).

  • Experience working in a highly regulated domain (ex. Healthcare, Banking, etc.

  • Strong knowledge of Structured Query Language (SQL) and Transact-SQL (T-SQL).

  • Experience using scripting languages such as JavaScript or Python.

  • Experience with agile delivery methodologies.

  • Strong organizational, administrative, and analytical skills required.

  • Travel may be required between 5 to 10%.

Preferred Requirements: .

  • Master’s degree in Data Science, Computer Science or MIS, Mathematics, Engineering, or related field.

  • DBT Certification.

  • Data Modeling experience preferred.

  • Experience working with Clinical Data and Regulatory Governance in a Healthcare setting.

  • Experience Healthcare data models, datasets, and source systems (e.g. EHR, claims, DICOM images, labs, etc.).

  • Experience with healthcare reference data (ICD, CPT etc.).

  • Experience with big data technologies (ex.Kafka, Hadoop, Spark) and containerization (ex. Docker, Kubernates).

  • Experience with FHIR and FHIR store formats and data repositories.

  • Hands-on design and configuration experience with cloud services Google (related to data storage and processing).

  • Data Governance Principles and enterprise frameworks (Warehousing, Data-as-a-Product, MESH, MDM).

  • Knowledge of HIPAA; ability to implement systems and processes in accordance with regulations.

#LI-remote

Benefits & conditions

The estimated annual pay range for this role is $125,000 - $150,000. Actual compensation depends on experience, skills, location, and other factors such as internal equity and budget.

Akumin is a leading provider of outpatient radiology and oncology services, partnering with top hospitals and health systems nationwide to deliver advanced diagnostic imaging and exceptional patient care close to home. With a national footprint, cutting-edge technology, and a strong commitment to innovation and patient-centered care, our teams play a vital role in improving outcomes for millions of patients each year.

About the company

Akumin Operating Corp. and its divisions are an equal opportunity employer and we believe in strength through diversity. All qualified applicants will receive consideration for employment without regard to, among other things, age, race, religion, color, national origin, sex, sexual orientation, gender identity & expression, status as a protected veteran, or disability.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:30 min

Scaling agile frameworks and data interoperability in healthcare

Leo Lindhorst · WWC 2022

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

Videos

See all

Related articles

See all