Principal Data Engineer (Databricks & Azure)

Syrencloud Llc
Atlanta, GA, United States
3 months ago
Apply on dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours
Job source

Tech stack

Airflow Application Frameworks Microsoft Azure Cloud Computing Code Review Information Systems Continuous Integration Data Validation Information Engineering Data Governance Data Infrastructure Extract Transform Load (ETL)
+18 more
Data Warehousing Apache Hive Python (Programming Language) Meta-Data Management Microsoft SQL Server Azure Data Lake SQL Databases SQL Server Integration Services Transact-SQL Cloud Platform System Azure Data Factory Git Data Lakes Pyspark Information Technology Data Lineage Data Pipelines Databricks

Job description

We are seeking a highly skilled Senior Data Engineer to join our growing Data Engineering team. This is a hands-on engineering role focused on building and delivering scalable data products within a modern cloud-based data platform using Databricks on Azure. The ideal candidate will have strong expertise in PySpark, Python, SQL, Delta Lake, and Azure Data Services, with experience developing production-grade data pipelines and supporting enterprise analytics initiatives.

This role offers the opportunity to work on cloud-first data engineering initiatives while supporting a strategic migration from legacy SQL Server data warehouse environments to a modern lakehouse architecture., Design, build, and maintain scalable data pipelines using Databricks, PySpark, Spark SQL, and Delta Lake

Develop and support data products following Medallion Architecture (Bronze/Silver/Gold)

Build and maintain SCD Type 1 & Type 2 data models and dimensional data warehouse solutions

Create and optimize ETL/ELT workflows using Azure Data Factory and Databricks Jobs

Implement data quality checks, monitoring, alerting, and operational support processes

Collaborate with business stakeholders across Finance, Supply Chain, Merchandising, Marketing, and Operations

Support legacy SQL Server and SSIS workloads during migration initiatives

Participate in code reviews, CI/CD deployments, testing, documentation, and production support

Contribute to engineering best practices, reusable frameworks, and technical mentoring

Requirements

5+ years of Data Engineering experience

2+ years working with cloud-based data platforms (Databricks preferred)

Strong experience with:

  • Databricks
  • PySpark
  • Spark SQL
  • Python
  • Delta Lake
  • Azure Data Factory
  • Azure Data Lake Storage Gen2
  • Azure DevOps
  • SQL Server / T-SQL

Expertise in:

  • Data Modeling
  • Star Schema Design
  • SCD Type 1 & Type 2
  • ETL/ELT Development
  • Data Warehousing
  • CI/CD Practices
  • Git Version Control

Experience building and supporting production-grade data pipelines

Strong problem-solving, communication, and stakeholder management skills, Databricks Data Engineer Certification

Experience with dbt

Familiarity with Apache Airflow

Retail, Supply Chain, Merchandising, or ERP domain experience

Experience with Data Governance, Metadata Management, and Data Lineage

Bachelor’’s Degree in Computer Science, Information Systems, or related field, * Ownership mentality and strong delivery focus

  • Ability to independently take requirements from design to production
  • Passion for cloud technologies and modern data platforms
  • Strong collaboration and communication skills
  • Experience working in compliance-sensitive environments (SOX, PCI, etc.)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all