Lead Data Engineer
Sovereign Technologies
Eagan, MN, United States
2 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours
Job source
Tech stack
Agile Methodology
Artificial Intelligence
Airflow
Amazon Web Services
Amazon Elastic Compute Cloud
Amazon S3
Big Data
Information Engineering
Extract Transform Load (ETL)
Database Queries
Distributed Computing Environment
Distributed Systems
+13 more
Healthcare Effectiveness Data and Information Set
Python (Programming Language)
Scrum Methodology
Standard Sql
SAS (Software)
Scripting
Freeform SQL
Generative AI
Git
Pyspark
Software Version Control
Data Pipelines
Databricks
Job description
We are seeking a Senior/Lead Data Engineer to modernize and own enterprise data pipelines by migrating legacy SAS-based workflows to Python/PySpark on Databricks. The engineer will partner with SMEs during the transition period and eventually take ownership of a critical healthcare analytics pipeline supporting HEDIS reporting., * Migrate legacy SAS pipelines to Python/PySpark on Databricks.
- Design, build, and maintain scalable ETL/ELT pipelines.
- Develop distributed data processing solutions using Databricks.
- Create and optimize complex SQL queries.
- Schedule, automate, and monitor data pipelines.
- Work with AWS services including S3, Lambda, Glue, and EC2.
- Manage Databricks notebooks, workflows, and clusters.
- Collaborate with business stakeholders and SMEs to transition pipeline ownership.
- Follow Agile methodologies and Git-based development practices.
Requirements
- 8+ years of Data Engineering experience.
- Strong Python programming and scripting skills.
- Hands-on PySpark development.
- Extensive Databricks experience (Notebooks, Workflows, Cluster Management).
- Experience processing large datasets in distributed environments.
- Strong SQL skills.
- AWS experience with:
- S3
- Glue
- Lambda
- EC2
- Experience building and automating ETL/ELT pipelines.
- Git/version control experience.
- Agile/Scrum experience.
Preferred Skills
- Experience with SAS-to-Python migration.
- AI/Automation experience.
- Healthcare or HEDIS domain experience.
- Lead/Principal Data Engineering experience.
Must-Have Skills
- Python
- PySpark
- Databricks
- SQL
- AWS (S3, Glue, Lambda, EC2)
- ETL/ELT Pipeline Development
- Distributed Data Processing
- Git
- Agile/Scrum
Nice-to-Have Skills
- SAS Modernization
- AI/Generative AI
- Healthcare/HEDIS
- Airflow or other workflow orchestration tools
- Leadership/Pipeline Ownership experience
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
EM
Eli McGarvie
about 3 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
DS
Dhannush Subramani
Top Big Data Technologies That You Need to Know
about 4 years ago
BB
Benedikt Bischof
Making Data Warehouses Fast: A Developer’s Story
about 4 years ago
EM
Eli McGarvie
Data Analyst Salary in the UK
about 3 years ago
IK
Igor Khokhriakov
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
24 days ago