Data Engineer - Remote

Unitedhealth Group Inc
Minnetonka, MN, United States
3 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Microsoft Azure Databases Data Validation Information Engineering Data Governance Extract Transform Load (ETL) Python (Programming Language) Query Optimization SQL Databases Data Streaming Data Processing
+12 more
Cloud Platform System Azure Data Factory Apache Spark Pyspark Data Lineage Low Latency Apache Kafka Spark Streaming Video Streaming Data Delivery Data Pipelines Databricks

Job description

Experteer Overview In this role you will design, build, and maintain scalable data pipelines on an Azure-based platform to support batch and real-time analytics. You will collaborate with data engineers, data science, and reporting partners to meet evolving data requirements and enforce governance, quality, and performance standards. The role focuses on end-to-end pipeline development, medallion architecture, and reliable data delivery at scale. You will work in a flexible, remote-friendly environment with opportunities to impact health care analytics and decision-making. Compensation / Benefits * Design, develop, and maintain batch and streaming data pipelines using Azure Data Factory and PySpark (Databricks) * Implement real-time ingestion with Spark Structured Streaming and schedule batch ETL jobs * Apply Medallion Architecture (Bronze/Silver/Gold) for data processing and data quality checks * Optimize Spark jobs, tune configurations, and manage resources for high throughput and low latency * Build and maintain ETL transformations, handle data from APIs, databases, file feeds, and IoT streams * Implement data quality validation, monitoring, and automated alerts * Utilize modern pipeline frameworks (Delta Live Tables, Lakehouse pipelines) where applicable * Document workflows, enforce data governance, and ensure data lineage and security controls Tasks * 3+ years in data engineering designing and implementing data pipelines and ETL * 2+ years SQL for data manipulation and query optimization * 1+ years Azure, Databricks or equivalent cloud data platform experience * Proficiency in Python and PySpark for batch and streaming transformations * Experience with streaming technologies (Spark Streaming, Kafka, Azure Event Hubs) * Strong collaboration in agile teams and ability to communicate with both technical and non-technical stakeholders * Experience with data quality validation and data governance practices Key requirements * comprehensive benefits package * equity stock purchase * 401k contribution * remote-friendly / flexible work * incentive and recognition programs

Requirements

latency * Build and maintain ETL transformations, handle data from APIs, databases, file feeds, and IoT streams * Implement data quality validation, monitoring, and automated alerts * Utilize modern pipeline frameworks (Delta Live Tables, Lakehouse pipelines) where applicable * Document workflows, enforce data governance, and ensure data lineage and security controls Tasks * 3+ years in data engineering designing and implementing data pipelines and ETL * 2+ years SQL for data manipulation and query optimization * 1+ years Azure, Databricks or equivalent cloud data platform experience * Proficiency in Python and PySpark for batch and streaming transformations * Experience with streaming technologies (Spark Streaming, Kafka, Azure Event Hubs) * Strong collaboration in agile teams and ability to communicate with both technical and non-technical stakeholders * Experience with data quality validation and data governance practices Key requirements * comprehensive benefits package * equity aa and purchase * 401k contribution * remote-friendly / flexible work * incentive and recognition programs

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:27 min

Managing traffic and tracking costs with Databricks Unity Catalog

Viktoria Semaan Viktoria Semaan · WWC Europe 2026

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:50 min

Executing LoRA fine-tuning using serverless Databricks AI runtimes

Viktoria Semaan Viktoria Semaan · WWC Europe 2026

Videos

See all

Related articles

See all