Sr Data Engineer - Data & Intelligence

Sumeru INC
Frisco, TX, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Microsoft Azure Big Data Cloud Storage Data Architecture Data Validation Data Transformation Data Warehousing Database Queries DevOps Dimensional Modeling Github Python (Programming Language)
+24 more
Key Management Log Analysis SQL Azure Query Optimization Role-Based Access Control Power BI Azure Active Directory Azure Data Lake SQL Databases Data Streaming Azure Data Factory Cloud Monitoring Delivery Pipeline Git Microsoft Fabric Data Lakes Pyspark Semi-structured Data Git Flow Star Schema Data Management Data Pipelines Databricks Programming Languages

Job description

This role builds, and maintains scalable data pipelines and lakehouse infrastructure on Microsoft Azure to support efficient extraction, transformation, and loading of data across batch and real-time workloads. It involves implementing and managing the Medallion Architecture (Bronze Silver Gold) using Azure Data Factory, Databricks-PySpark, and Azure SQL Database and Databricks unity catalogue. The role requires ensuring SLA-adherent data quality standards. Success is measured by pipeline reliability, data freshness SLA compliance, and the quality of Gold-layer datasets powering Power BI executive dashboards. The work supports organizational decision-making by delivering trusted, well-governed data to business executives and analytics consumers.

Requirements

Experience building and optimizing big data pipelines using Azure Data Factory, PySpark, and SQL across structured and semi-structured data sets Hands-on experience implementing Medallion Architecture (Bronze/Silver/Gold) Experience with Delta Lake - ACID transactions, incremental loading, schema evolution, partitioning strategies Experience performing root cause analysis on pipeline failures and data quality issues to resolve SLA breaches and identify platform improvement opportunities Azure Foundational Services : Working knowledge of: Azure Data Factory (ADF), ADLS Gen2, Azure SQL Database, Azure Blob Storage, Azure Key Vault, Azure Monitor / Log Analytics, Azure Event Hubs, Microsoft Fabric Lakehouse, Azure Active Directory / Entra ID (RBAC, Service Principals) Programming Languages: Proficiency in Python and PySpark for data transformation, pipeline automation, and large-scale distributed processing; strong SQL skills including window functions, CTEs, and query optimization across relational and lakehouse engines Data Architecture: Solid understanding of Medallion Architecture, dimensional modeling (Star Schema, SCD Types 1/2/3), and the trade-offs between lakehouse, data warehouse, and data lake patterns Pipeline Engineering: Ability to build robust ADF pipelines with ForEach, Lookup, Copy Activity, and Data Flows; incremental loading via watermark or CDC; error handling, retry logic, and dead-letter patterns Data Quality Experience: Experience implementing SLA-based data quality checks (freshness, completeness, row count), monitoring via Azure Monitor and ADF diagnostic logs, and defining data quality agreements with business stakeholders. DevOps for Data: Experience with Git-based workflows, ADF Git integration, CI/CD pipeline promotion across Dev/Test/Prod using Azure DevOps or GitHub Actions Reporting Layer Awareness: Understanding of how Gold-layer data feeds Power BI - DirectQuery vs. Import mode trade-offs, dataset refresh patterns, and semantic model collaboration with BI teams Ability to manage work across multiple concurrent pipeline projects, prioritize by business impact, and communicate status clearly to technical and non-technical stakeholders Good to have skills: Experience with Microsoft Fabric (Lakehouse, Notebooks, OneLake, Fabric Pipelines) - active migration or greenfield project Experience with real-time / streaming workloads using Azure Event Hubs or Structured Streaming in PySpark Experience delivering data platforms for executive-level reporting via Power BI semantic

About the company

Jones Lang LaSalle

  • Allen, TX JLL empowers you to shape a brighter way. Our people at JLL are shaping the future of real estate for a better world by combining world class services, advisory and technology fo…, Jones Lang LaSalle

  • Carrollton, TX JLL empowers you to shape a brighter way. Our people at JLL are shaping the future of real estate for a better world by combining world class services, advisory and technology fo…, Jones Lang LaSalle

  • Richardson, TX JLL empowers you to shape a brighter way. Our people at JLL are shaping the future of real estate for a better world by combining world class services, advisory and technology fo…

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all