Senior Data Engineer - Healthcare Data & Audience Applications

Zeta Global
Nashville, TN, United States
27 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$140,000.0 - $160,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Airflow Amazon S3 Batch Processing Big Data Code Review Continuous Integration Data as a Services Data Architecture Information Engineering Data Warehousing Apache Hive
+14 more
Python (Programming Language) Meta-Data Management Standard Sql SQL Databases Tokenization Data Storage Technologies Delivery Pipeline Snowflake Electronic Medical Records Kubernetes Data Pipelines User Identification Amazon Elastic Mapreduce (EMR) Docker

Job description

Zeta Global is seeking a Senior Data Engineer to build reliable, scalable data pipelines and data products for a healthcare vertical. You will be a hands-on engineer who turns complex healthcare and marketing datasets into trusted foundations for audience discovery, segmentation, activation, reporting, and measurement.

Working closely with the engineering and product team, you will help implement the team’s data architecture and engineering standards while owning significant parts of the delivery lifecycle. You will contribute to well-designed, production-ready systems-not define the overall architecture or technical roadmap alone., * Design, develop, test, deploy, and operate production-grade pipelines for healthcare, identity, audience, media-exposure, and campaign-performance data using Python, SQL, Airflow, S3, Snowflake, and EMR.

  • Implement maintainable data models, transformations, governed views, and reusable datasets for provider identity, claims/Rx, NPI/HCP, media, brand, and connector data.
  • Deliver data products that support HCP and patient/DTC audience discovery, segmentation, activation, measurement, and reporting.
  • Build Airflow workflows with clear dependencies, retries, alerting, data-quality checks, and operational runbooks; use EMR for large-scale enrichment, normalization, and other compute-intensive workloads.
  • Write efficient SQL across Snowflake, Hive, and Athena, adapting to platform-specific syntax and query behavior.
  • Partner with product, analytics, data science, and platform teams to translate business and healthcare requirements into resilient technical solutions.
  • Implement data-quality controls, reconciliation checks, monitoring, alerting, and incident-response practices for critical data products.
  • Support data onboarding and integration for healthcare partners and internal sources, including validation, normalization, and source-to-target mapping.
  • Apply privacy-by-design practices for PHI/PII, including access controls, masking, approved joins, retention, and auditability.
  • Collaborate with the Lead Data Engineer on technical designs, code reviews, documentation, and delivery plans; mentor less-experienced engineers as needed.
  • Troubleshoot production issues and improve pipeline performance, reliability, and observability over time.

Core Technical Environment

  • Data storage & warehouse: Snowflake Native and Amazon S3.
  • Orchestration: Apache Airflow for general pipeline setup and scheduling.
  • Heavy processing: Amazon EMR for targeted, compute-intensive jobs.
  • Programming: Python for Airflow pipelines and supporting data engineering services.
  • Querying: SQL in Snowflake, Hive, and Athena.

Requirements

  • 5-8 years of hands-on data engineering experience, including ownership of production pipelines and data models, with experience working with healthcare data such as provider/HCP, claims, prescription, patient/DTC, or healthcare audience datasets.
  • Strong Python and expert SQL skills, with demonstrated experience building transformations, optimizing queries, and diagnosing data issues.
  • Hands-on experience with AWS data services, especially S3, and a modern cloud data warehouse; experience with Snowflake, Airflow, and EMR is strongly preferred.
  • Experience with data modeling, schema evolution, batch processing, orchestration, testing, CI/CD, and production support practices.
  • Proven ability to work with large, complex datasets and deliver reliable, well-documented data products.
  • Deep, practical knowledge of HIPAA, PHI/PII handling, privacy-by-design controls, and the operational requirements of regulated healthcare data environments.
  • Experience with AdTech/MarTech, identity resolution, audience onboarding, segmentation, data linkage, media measurement, attribution, or campaign reporting.
  • Ability to balance healthcare privacy constraints with the need for timely, accurate audience and performance insights.
  • Strong collaboration and communication skills across engineering, product, analytics, and business stakeholders.

Preferred

  • Experience with healthcare data providers, identity ecosystems, tokenization, clean rooms, or privacy-enhancing technologies.
  • Experience with data cataloging, lineage, observability, and data-quality frameworks.
  • Experience with Docker, Kubernetes/EKS, infrastructure as code, and cloud deployment workflows.
  • Experience supporting reporting, attribution, or measurement products tied to campaign or business outcomes.
  • Exposure to ML/AI-enabled data products or analytics workflows.

Benefits & conditions

2.72.7 out of 5 stars Nashville, TN Remote $140,000 - $160,000 a year, Pulled from the full job description

  • Pet insurance
  • Health insurance
  • Employee discount
  • Vision insurance
  • Dental insurance
  • Unlimited paid time off, * Unlimited PTO
  • Excellent medical, dental, and vision coverage
  • Employee Equity
  • Employee Discounts, Virtual Wellness Classes, and Pet Insurance And more!!

About the company

Zeta Global (NYSE: ZETA) is the AI-Powered Marketing Cloud that leverages advanced artificial intelligence (AI) and trillions of consumer signals to make it easier for marketers to acquire, grow, and retain customers more efficiently. Through the Zeta Marketing Platform (ZMP), our vision is to make sophisticated marketing simple by unifying identity, intelligence, and omnichannel activation into a single platform - powered by one of the industry’s largest proprietary databases and AI. Our enterprise customers across multiple verticals are empowered to personalize experiences with consumers at an individual level across every channel, delivering better results for marketing programs. Zeta was founded in 2007 by David A. Steinberg and John Sculley and is headquartered in New York City with offices around the world. To learn more, go to www.zetaglobal.com., We’re committed to building a workplace culture of trust and belonging, so everyone feels invited to bring their whole selves to work. We provide a forum for employees to celebrate, support and advocate for one another. Learn more about our commitment to diversity, equity and inclusion here: https://zetaglobal.com/blog/a-look-into-zetas-ergs

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all