Data Engineer

Arcadia Inc.
Sandy Springs, GA, United States
5 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
4 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Microsoft Azure Continuous Integration Information Engineering Data Infrastructure Python (Programming Language) Power BI SQL Databases Apache Spark Data Layers Microsoft Fabric Pyspark
+3 more
Azure Synapse Analytics Data Pipelines Databricks

Job description

Seeking a Data Engineer to help build and launch a new Azure-based lakehouse environment. This role focuses on designing scalable data pipelines, implementing lakehouse architecture, and delivering clean, governed datasets for reporting, analytics, and AI/ML.

You’ll work across Azure and Microsoft Fabric (OneLake, Data Factory, Power BI, Databricks, Synapse) to create ingestion patterns, curated data models, semantic layers, and AI-ready datasets. Responsibilities include ensuring data quality, lineage, security, and governance; optimizing pipeline performance; automating workflows with CI/CD; and partnering closely with business, analytics, and AI teams.

Requirements

Ideal candidates have 4+ years of hands-on data engineering experience, strong Azure/Fabric expertise, Spark/PySpark/SQL/Python skills, and a track record of contributing to new data platform or lakehouse builds. Experience preparing governed datasets for AI/ML and building dimensional/semantic models for Power BI is highly valued. Strong communication skills and comfort working in broad-ownership or small-team environments are preferred.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

1:24 min

Moving the semantic layer upstream to avoid vendor lock-in

Piotr Menclewicz Piotr Menclewicz · Europe 2026 Virtual

2:27 min

Managing traffic and tracking costs with Databricks Unity Catalog

Viktoria Semaan Viktoria Semaan · World Congress 2026 Europe

1:59 min

Evolving roles in AI driven software teams

Ignacio Riesgo Ignacio Riesgo +1 · World Congress 2024

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all