Senior Data Engineer - Databricks

Data Canopy Colocation LLC
Raleigh, NC, United States
10 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$73,450.0 - $132,775.0
Working hours
Regular working hours

Tech stack

Unity 3d Airflow Amazon Web Services Microsoft Azure Cloud Computing Computer Programming Continuous Integration Information Engineering Data Governance Extract Transform Load (ETL) Data Security Distributed Systems
+14 more
Apache Hive Python (Programming Language) Meta-Data Management Performance Tuning SQL Databases Data Streaming Cloud Platform System Snowflake Apache Spark Git Data Lakes Pyspark Data Pipelines Databricks

Requirements

  • Build scalable, production-grade ETL/ELT pipelines using Databricks (PySpark, Spark SQL, Delta Live Tables, Workflows).

  • Ingest structured, semi-structured, and streaming data into Bronze, Silver, and Gold layers.

  • Develop optimized transformations, data quality rules, and reusable framework components.

  • Implement best practices for job orchestration, monitoring, alerting, and automation.

  • Hands-on experience: Spark, Delta Lake, Workflows, Unity Catalog.

  • Strong SQL programming and performance tuning skills.

  • Experience with cloud environments (AWS/Azure/GCP).

  • Experience with modern data lakehouse concepts and distributed systems.

  • Strong understanding of Lakeflow Connect, LSDP/Lakehouse, Medallion Architecture, Data Validations, Genie, and Agent Bricks/RAG use cases.

  • Should be able to explain these concepts using real project examples and architecture decisions.

Requirements

  • Strong Python (PySpark) and SQL programming

  • Databricks - Spark, Delta Lake, Workflows, Unity Catalog

  • ETL/ELT pipeline development - Medallion Architecture (Bronze/Silver/Gold)

  • Delta Live Tables, Auto-Loader, Structured Streaming

  • Data modeling - dimensional (star/snowflake), normalization/denormalization

  • CI/CD, Git, job orchestration

  • Cloud experience - AWS, Azure, or GCP

  • 7-10+ years in data engineering

Nice-to-Have Skills

  • Lakeflow Connect, LSDP/Lakehouse, Genie, Agent Bricks/RAG use cases

  • Data governance, metadata management, Unity Catalog advanced features

  • Airflow, dbt, or similar orchestration tools

  • Data security, compliance, and access models

  • Cost optimization and performance tuning in cloud environments

  • Corporate/enterprise data warehousing background

About the company

DATAECONOMY is one of the fastest-growing Data & Analytics company with global presence. We are well-differentiated and are known for our Thought leadership, out-of-the-box products, cutting-edge solutions, accelerators, innovative use cases, and cost-effective service offerings. We offer products and solutions in Cloud, Data Engineering, Data Governance, AI/ML, DevOps and Blockchain to large corporates across the globe. Strategic Partners with AWS, Collibra, cloudera, neo4j, DataRobot, Global IDs, tableau, MuleSoft and Talend., DATAECONOMY

  • Raleigh, NC DATAECONOMY is one of the fastest-growing Data & Analytics company with global presence. We are well-differentiated and are known for our Thought leadership, out-of-the-box product…

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all