Data Engineer

Empower Professionals
United States
3 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Working hours
Regular working hours
Job source

Tech stack

Business Logic Cloud Computing Data Architecture Data Infrastructure Data Integrity Data Transformation Data Structures Apache Hive Metadata Query Optimization SQL Stored Procedures SQL Databases
+8 more
Data Streaming Transact-SQL Enterprise Data Management Build Management Microsoft Fabric Pyspark Data Management Data Pipelines

Job description

We are seeking a Data Engineer to support a Microsoft Fabric modernization initiative focused on migrating existing on-premises SQL Stored Procedures workloads into a Fabric Lakehouse environment. This role will be responsible for converting SQL/T-SQL stored procedures into Fabric notebooks, building and orchestrating data pipelines, validating migrated datasets, and creating technical documentation to support ongoing operations. This position will primarily support the Silver Layer within a Medallion Architecture and will play a critical role in ensuring data accuracy, consistency, and scalability throughout the migration effort. Team structure is small, must be highly motivated and self-motivated. Will have access to some SMEs but this role is primarily responsible for the vast majority of the work here. Must be comfortable working in a small team with one other to deliver fast and quick with great quality., Convert existing SQL/T-SQL stored procedures into Microsoft Fabric Notebooks using PySpark and Spark SQL.

Design and build the Silver Layer of a medallion architecture in Fabric using Lakehouse, ensuring parity between Fabric data structures and existing on-premises SQL environments.

Develop and maintain orchestration frameworks using Microsoft Fabric Data Pipelines.

Build automated validation and reconciliation processes to ensure migrated data matches source systems on a 1:1 basis.

Create reusable testing frameworks and validation artifacts to continuously verify data integrity and accuracy.

Support data pipeline monitoring, scheduling, and orchestration activities within Microsoft Fabric.

Partner with architects, engineers, and business stakeholders to ensure successful migration outcomes.

Develop and maintain technical documentation detailing solution design, implementation approaches, validation procedures, and operational support processes.

Analyze and troubleshoot data issues across Fabric notebooks, pipelines, and Lakehouse objects.

Requirements

5-7+ years of experience as a Data Engineer or similar role.

1-2+ years of hands-on experience with Microsoft Fabric.

Strong expertise in SQL and T-SQL, including:

Stored procedures

Query optimization

Data transformations

SQL-based business logic

Experience migrating or translating SQL-based processing into modern cloud-based data platforms.

Strong hands-on experience with:

PySpark

Spark SQL

Microsoft Fabric Notebooks

Experience developing and managing data pipelines and orchestration workflows.

Deep understanding of:

Lakehouse architecture

Microsoft Fabric data flows

Data Pipelines

Medallion Architecture concepts

Experience building automated data quality, reconciliation, and validation frameworks.

Strong technical documentation and communication skills.

Preferred Qualifications

Experience working with metadata-driven frameworks.

Experience implementing orchestration and scheduling strategies within Microsoft Fabric.

Familiarity with enterprise data migration initiatives.

Experience supporting cloud-based analytics and modern data platform solutions.

Technical Environment

Microsoft Fabric

Fabric Notebooks

Fabric Data Pipelines

Lakehouse Architecture

Medallion Architecture (Silver Layer Focus)

SQL / T-SQL

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

1:47 min

Comparing Egeria to alternative open metadata solutions

Ferd Scheepers · World Congress 2022

1:19 min

The true role and evolution of data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

2:08 min

Creating standard APIs via the Egeria open metadata project

Ferd Scheepers · World Congress 2022

Videos

See all

Related articles

See all