Senior Data Engineer

InnoVet Health LLC
United States
14 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Compensation
$130,000.0
Working hours
Regular working hours
Job source

Tech stack

Cerner Artificial Intelligence ASC X12 Standards Microsoft Azure Health Informatics Data Architecture Information Engineering Data Integration Extract Transform Load (ETL) Data Security Distributed Computing Environment Interoperability
+19 more
Performance Tuning Role-Based Access Control Azure Data Lake SQL Databases Data Streaming Azure Service Bus Azure Data Factory Fast Healthcare Interoperability Resources Electronic Medical Records Data Lakes Pyspark Information Technology Data Analytics Health Level Seven International Stream Analytics Data Pipelines Serverless Computing Key Vault Databricks

Job description

InnoVet Health, a small and growing business that provides health IT professional services to the Department of Veterans Affairs (VA) is looking for an experienced Sr. Azure Data Engineer proficient in Databricks, Delta Lake, and medallion/lakehouse architectures to support healthcare analytics, interoperability, and data integration needs. You will build scalable data pipelines, ensure data quality, and support advanced analytics across clinical, operational, and regulatory datasets. Your work will directly impact VA healthcare delivery by building an analytical infrastructure that enables AI/ML and advanced analytics to deliver short-term and longitudinal insights to healthcare professionals serving Veterans. You will work in an agile environment alongside VA and contractor stakeholders. It is flexible, full-time, and does not require relocation (work-from-home). The pay, benefits, and growth potential are competitive.

Responsibilities

  • Gather and translate business, technical, and functional requirements into data architecture and pipeline design decisions. Design and develop Azure Data Factory and Databricks-based ETL/ELT pipelines using PySpark, Delta Lake, and medallion/lakehouse architecture.
  • Ingest and transform healthcare data (clinical, claims, FHIR, HL7, EHR, ADT, PGHD) from diverse sources.
  • Build secure, scalable solutions using Azure Data Lake Storage, ADF, and Event Hubs, and related services, with attention to latency and reliability requirements.
  • Implement data quality, lineage, and governance using Microsoft Purview.
  • Optimize Databricks jobs (performance tuning, cluster sizing, Z-ordering, partitioning).
  • Enforce HIPAA-aligned security practices: RBAC, Key Vault, private endpoints, PHI protection.
  • Collaborate with data scientists, analysts, and clinical informatics teams.
  • Stay up to date with emerging technologies and trends in data engineering and healthcare data management.
  • Present and discuss results with IT and business stakeholders.
  • Participate in company growth and other responsibilities, as assigned.

Requirements

  • Bachelor’s or master’s degree in computer science, data analytics, or related field.
  • Minimum 6+ years data engineering experience; 4+ years hands-on with Azure and 2+ years hands-on with Databricks.
  • Strong skills in PySpark, Delta Lake, SQL, and distributed data processing.
  • Experience with healthcare data standards (FHIR, HL7, X12/EDI, CCD, claims data, PGHD).
  • Strong understanding of HIPAA, PHI handling, and secure data architecture.
  • Experience with ADF, ADLS Gen2, Azure Functions, and event-driven ingestion.
  • Strong understanding of data modeling for analytics (dimensional + lakehouse).
  • Excellent problem-solving, collaboration and communication skills.
  • Green card or US citizen required because of government contract work.
  • No 1099 or corp-to-corp or international outsourcing or staffing agencies.

Preferred

  • Experience with Federal EHR (VistA and Oracle Health) data.
  • Experience with Azure Event Hubs, Stream Analytics, AWS Kinesis, or similar data streaming platforms is also a plus., * This position works directly with US government contracts and under Order 11935, it requires either U.S. Citizenship or valid Green Card. Please answer 2 if you are a US citizen, 1 if you have a valid green card, 0 if neither. Please make sure to answer this question.
  • Your LinkedIn profile link (required):

Education:

  • Bachelor’s (Required)

Experience:

  • Azure Data Lake: 4 years (Required)
  • Databricks: 4 years (Required)

Benefits & conditions

United States Remote From $130,000 a year - Full-time, Pulled from the full job description

  • Referral program
  • 401(k)
  • Health insurance
  • 401(k) matching
  • Paid time off
  • Vision insurance
  • Dental insurance, * 401(k)
  • 401(k) matching
  • Dental insurance
  • Health insurance
  • Paid time off
  • Referral program
  • Vision insurance

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

4:55 min

Centralizing credentials management using Azure Key Vault resource integration

Marcel Lupo · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:21 min

Structuring the backend architecture of machine learning workspaces

Jose Luis Latorre Millas · LIVE

Videos

See all

Related articles

See all