Azure Databricks Developer

Kainos
Houston, TX, United States
2 months ago
Apply on dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Airflow Microsoft Azure Big Data Computer Programming Data Architecture Data Integration Extract Transform Load (ETL) Data Systems Python (Programming Language) Azure Data Factory Apache Spark Data Lakes
+3 more
Pyspark Data Pipelines Databricks

Job description

We are looking for an experienced Azure Databricks Developer with strong hands-on skills in building scalable data pipelines on Azure. The ideal candidate should have solid experience with Databricks, Spark, and Azure data services., * Develop and maintain data pipelines using Azure Databricks

  • Process large datasets using PySpark / Apache Spark
  • Design and optimize ETL workflows
  • Work with Azure data services for data integration and storage
  • Ensure performance, scalability, and reliability of data solutions
  • Collaborate with cross-functional teams to deliver data projects

Requirements

  • Strong experience in Azure Databricks
  • Hands-on experience with PySpark / Apache Spark
  • Good programming skills in Python or Scala
  • Experience in ETL and data pipeline development
  • Working knowledge of Azure services like Data Factory (ADF), ADLS, etc.

Preferred Skills:

  • Experience with Delta Lake / Lakehouse architecture
  • Knowledge of Airflow or other orchestration tools
  • Exposure to Azure cloud architecture

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

1:43 min

AWS infrastructure stack and data flow pipeline overview

Artem Volk Artem Volk +1 · World Congress 2024

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

Videos

See all

Related articles

See all