> Markdown version of [/jobs/ext/1487535-data-engineer](https://www.wearedevelopers.com/jobs/ext/1487535-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Page Michael International Inc - **Location:** New York, NY, United States - **Experience:** Experienced - **Salary:** $135,000.0 - **Contract:** Permanent contract - **Skills:** Query Performance, Automation of Tests, Cloud Computing, Code Review, Information Systems, Data Dictionary, Information Engineering, Extract Transform Load (ETL), Data Systems, Data Warehousing, Distributed Computing Environment, Distributed Data Store, Python (Programming Language), Metadata, Operational Databases, Oracle (Applications), Performance Tuning, Release Management, SQL Databases, Enterprise Data Management, Software Organization, Data Processing, Warehouse Management Systems, Apache Spark, Git, Pyspark, Storage Technologies, Information Technology, People Soft, Software Version Control, Data Pipelines, Databricks - **Published:** July 29, 2026 - **Apply:** https://www.michaelpage.com/job-detail/data-engineer/ref/jn-072026-7072562 ## About the Role * Bachelor's degree in computer science, data engineering, information systems, engineering, or a related field; Masters preferred * 3-5+ years of experience in data science/engineering, data warehousing, or analytics engineering. * Advanced programming skills in Python and strong proficiency in SQL for building and maintaining production data pipelines across development and production environments. * Experience developing scalable ELT/ETL pipelines and analytical or dimensional data models on relational, cloud, or distributed data processing platforms. * Experience with Git, code review, automated testing, and modern software development practices. * Strong understanding of data quality, troubleshooting, performance optimization, and production support. * Experience integrating complex enterprise data across multiple source systems. ## Description The Data Engineer will play a key role in designing, building, and maintainig scalable data pipelines, curated analytical data models, and cloud-based data products that support enterprise supply chain analytics and operations. * Data Engineer within healthcare * Hands on with Python and SQL - Unable to sponsor candidates, * Design, build, and maintain scalable batch data pipelines, transformation workflows, dimensional models, and curated analytical data models using SQL, Python, Databricks, Apache Spark/PySpark, and distributed data processing technologies. * Define and maintain source-to-target mappings and transformation logic for data integrated from enterprise platforms including Epic, Oracle, ParEx, GHX, PeopleSoft, and other supply chain systems. * Implement automated data quality, reconciliation, validation, and testing processes to ensure reliable, production-ready datasets. * Optimize data processing, query performance, compute utilization, and storage design across relational and distributed data platforms. * Develop, test, deploy, monitor, and support data solutions across development and production environments, following established release management, change management, and deployment practices. * Provide operational support for production data pipelines, including incident triage, root cause analysis, issue resolution, and recovery of failed workflows. * Implement workflow orchestration, data observability, monitoring, and alerting to ensure pipeline reliability, data freshness, data quality, and timely issue detection. * Maintain technical documentation, metadata, lineage, schema management, and data dictionaries to support governance, transparency, and operational support. * Partner with business and technical stakeholders to translate requirements into scalable, sustainable data engineering solutions. * Contribute to engineering best practices through code review, automated testing, version control, release management, and deployment standards. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [A Data Mesh needs Open Metadata](https://www.wearedevelopers.com/videos/505-a-data-mesh-needs-open-metadata) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries)