Data Engineer - Remote

Unitedhealth Group Inc
Minnetonka, MN, United States
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Agile Methodology Microsoft Azure Databases Data Validation Data Cleansing Information Engineering Data Governance Extract Transform Load (ETL) Document-Oriented Databases Python (Programming Language) Query Optimization
+15 more
SQL Databases Data Streaming Parquet Data Processing Cloud Platform System Azure Data Factory Apache Spark Data Lakes Pyspark Low Latency Apache Kafka Spark Streaming Video Streaming Data Pipelines Databricks

Job description

Experteer Overview As a Data Engineer, you design, build, and maintain scalable data pipelines on an Azure-based platform to enable batch and real-time analytics. You will collaborate with data engineers, data scientists, and reporting partners to meet evolving data needs and enforce governance, quality, and security across the data lifecycle. You’ll apply Medallion architecture (Bronze/Silver/Gold) and optimize Spark-based jobs for low latency and high throughput. This role offers remote work flexibility with in-office requirements in certain locations, and opportunities to impact large-scale health analytics. Compensation / Benefits * Design, develop, and maintain robust data pipelines for streaming and batch data using Azure Data Factory and PySpark via Databricks * Implement Medallion Architecture (Bronze/Silver/Gold) with data cleansing, transformation, and curated datasets * Optimize Spark jobs for performance, tune configurations, and scale clusters * Develop and maintain ETL processes; handle data from APIs, databases, files, and IoT streams * Implement data quality checks, automated alerts, and data governance practices * Leverage modern pipeline frameworks and declarative pipelines (e.g., Delta Live Tables, Lakehouse) * Monitor pipeline health, diagnose failures, and collaborate with teams to adjust pipelines * Document data workflows, lineage, and governance controls to ensure security and privacy Tasks * 3+ years in data engineering (designing and implementing data pipelines and ETL) * 2+ years in SQL for data manipulation and query optimization * 1+ years in Azure/Databricks or equivalent cloud data platform; experience with Delta Lake or Parquet * Proficiency in Python and Apache Spark (PySpark) for batch and streaming * Experience with streaming technologies (Spark Streaming, Kafka, Azure Event Hubs) * Agile team experience; strong communication with technical and non-technical stakeholders * Familiarity with data governance, security, and documentation practices * Ability to diagnose data issues and implement validation checks Key requirements * comprehensive benefits package * incentive and recognition programs * equity stock purchase * 401k contribution * remote work flexibility * telecommuter policy (for remote roles)

Requirements

security handle data from APIs, databases, files, and IoT streams * Implement data quality checks, automated alerts, and data governance practices * Leverage modern pipeline frameworks and declarative pipelines (e.g., Delta Live Tables, Lakehouse) * Monitor pipeline health, diagnose failures, and collaborate with teams to adjust pipelines * Document data workflows, lineage, and governance controls to ensure security and privacy Tasks * 3+ years in data engineering (designing and implementing data pipelines and ETL) * 2+ years in SQL for data manipulation and query optimization * 1+ years in Azure/Databricks or equivalent cloud data platform; experience with Delta Lake or Parquet * Proficiency in Python and Apache Spark (PySpark) for batch and streaming * Experience with streaming technologies (Spark Streaming, Kafka, Azure Event Hubs) * Agile team experience; strong communication with technical and non-technical stakeholders * Familiarity with data governance, security, and documentation practices * Ability to diagnose data issues and implement validation checks Key requirements * comprehensive benefits package * incentive and recognition programs * equity stock purchase * 401k contribution * remote work flexibility * telecommuter policy (for remote roles)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:50 min

How Parquet metadata enables efficient data reading

Matthias Niehoff Matthias Niehoff · WWC Europe 2026

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all