Principal Data Engineer

Robert Half
Avon, MN, United States
24 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Compensation
$130,000.0 - $190,000.0
Working hours
Regular working hours
Job source

Tech stack

Clean Code Principles Big Data Cloud Database Information Engineering Data Integration Extract Transform Load (ETL) Data Systems Software Design Patterns IT Management Python (Programming Language) Azure Data Lake SQL Databases
+10 more
Enterprise Software Applications Azure Data Factory Apache Spark Data Layers Microsoft Fabric Data Lakes Pyspark Information Technology Data Pipelines Databricks

Job description

  • Lead the design and ongoing evolution of a Microsoft Fabric lakehouse environment, establishing a scalable architecture across raw, refined, and curated data layers.
  • Create and support robust data integration workflows that bring information together from more than 30 enterprise applications and operational platforms.
  • Set technical standards for data engineering, including coding practices, solution design patterns, documentation, and review processes.
  • Mentor data engineers through hands-on coaching, collaborative development, and detailed code feedback to strengthen team capability.
  • Partner with data and IT leadership to align platform architecture with business priorities and future expansion needs.
  • Develop high-quality data pipelines using modern tools and frameworks such as SQL, Python, Spark, and PySpark within cloud-based data ecosystems.
  • Evaluate new technologies, architectural approaches, and data platform trends, then provide informed recommendations for adoption.
  • Support governance and platform reliability by promoting strong lineage, quality, performance, and scalable engineering practices.
  • Contribute to cross-platform interoperability initiatives, including environments that may involve Databricks, governance tooling, and enterprise data services.

Requirements

  • Bachelor’s degree in Computer Science, Data Science, Mathematics, Engineering, Business Analytics, or a related discipline; comparable relevant experience may be considered.
  • At least 10 years of progressive data engineering experience with increasing technical leadership and architectural responsibility.
  • Advanced expertise with Microsoft Fabric, including lakehouse design, pipelines, dataflows, notebooks, semantic models, and related platform capabilities.
  • Strong command of modern cloud data technologies such as Azure Data Factory, Azure Data Lake, Delta Lake, and enterprise-scale ETL and ELT frameworks.
  • Expert-level proficiency in SQL and Python, along with practical experience building solutions with Spark or PySpark.
  • Demonstrated success designing and implementing medallion-style lakehouse architectures or similar large-scale data platform models.
  • Experience integrating enterprise systems such as ERP, HR, CRM, service management, or field operations tools into a centralized data environment.
  • Familiarity with Databricks, Unity Catalog, Microsoft Purview, or relevant Microsoft and Azure data certifications is valued.

About the company

Robert Half is the world’s first and largest specialized talent solutions firm that connects highly qualified job seekers to opportunities at great companies. We offer contract, temporary and permanent placement solutions for finance and accounting, technology, marketing and creative, legal, and administrative and customer support roles.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

5:14 min

Executing Databricks jobs with built-in Airflow operators

Alan Mazankiewicz · LIVE

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all