Data Engineer

Page Michael International Inc
New York, NY, United States
14 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$135,000.0
Working hours
Regular working hours

Tech stack

Query Performance Automation of Tests Cloud Computing Code Review Information Systems Data Dictionary Information Engineering Extract Transform Load (ETL) Data Systems Data Warehousing Distributed Computing Environment Distributed Data Store
+20 more
Python (Programming Language) Metadata Operational Databases Oracle (Applications) Performance Tuning Release Management SQL Databases Enterprise Data Management Software Organization Data Processing Warehouse Management Systems Apache Spark Git Pyspark Storage Technologies Information Technology People Soft Software Version Control Data Pipelines Databricks

Job description

The Data Engineer will play a key role in designing, building, and maintainig scalable data pipelines, curated analytical data models, and cloud-based data products that support enterprise supply chain analytics and operations.

  • Data Engineer within healthcare
  • Hands on with Python and SQL - Unable to sponsor candidates, * Design, build, and maintain scalable batch data pipelines, transformation workflows, dimensional models, and curated analytical data models using SQL, Python, Databricks, Apache Spark/PySpark, and distributed data processing technologies.
  • Define and maintain source-to-target mappings and transformation logic for data integrated from enterprise platforms including Epic, Oracle, ParEx, GHX, PeopleSoft, and other supply chain systems.
  • Implement automated data quality, reconciliation, validation, and testing processes to ensure reliable, production-ready datasets.
  • Optimize data processing, query performance, compute utilization, and storage design across relational and distributed data platforms.
  • Develop, test, deploy, monitor, and support data solutions across development and production environments, following established release management, change management, and deployment practices.
  • Provide operational support for production data pipelines, including incident triage, root cause analysis, issue resolution, and recovery of failed workflows.
  • Implement workflow orchestration, data observability, monitoring, and alerting to ensure pipeline reliability, data freshness, data quality, and timely issue detection.
  • Maintain technical documentation, metadata, lineage, schema management, and data dictionaries to support governance, transparency, and operational support.
  • Partner with business and technical stakeholders to translate requirements into scalable, sustainable data engineering solutions.
  • Contribute to engineering best practices through code review, automated testing, version control, release management, and deployment standards.

Requirements

  • Bachelor’s degree in computer science, data engineering, information systems, engineering, or a related field; Masters preferred
  • 3-5+ years of experience in data science/engineering, data warehousing, or analytics engineering.
  • Advanced programming skills in Python and strong proficiency in SQL for building and maintaining production data pipelines across development and production environments.
  • Experience developing scalable ELT/ETL pipelines and analytical or dimensional data models on relational, cloud, or distributed data processing platforms.
  • Experience with Git, code review, automated testing, and modern software development practices.
  • Strong understanding of data quality, troubleshooting, performance optimization, and production support.
  • Experience integrating complex enterprise data across multiple source systems.

Benefits & conditions

New York, New York Permanent USD115,000 - USD135,000 per year View Job Description

About the company

One of the nation’s top integrated academic health systems devoted to patient care, education, and research.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.michaelpage.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

1:47 min

Comparing Egeria to alternative open metadata solutions

Ferd Scheepers · WWC 2022

1:19 min

The true role and evolution of data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all