Databricks Engineer - REMOTE

Databricks
United States
2 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Amazon Web Services Amazon S3 Application Frameworks Microsoft Azure Business Intelligence Development BigQuery Cloud Computing Cloud Database Databases Data Validation
+44 more
Information Engineering Data Governance Extract Transform Load (ETL) Data Systems Data Warehousing IBM DB2 DevOps Github Apache Hive Identity and Access Management Python (Programming Language) Key Management Mainframes Meta-Data Management Microsoft SQL Server Oracle (Applications) Performance Tuning Standard Sql Azure Data Lake SAP (Applications) SAP NetWeaver Business Warehouse Data Streaming Systems Integration Teradata SQL Enterprise Data Management Feature Engineering Data Ingestion Azure Data Factory Snowflake Git Event Driven Architecture Data Lakes Pyspark Data Lineage Apache Kafka Spark Streaming Machine Learning Operations Cloud Integration Restful APIs Stream Processing Data Pipelines TIBCO (Software) Legacy Systems Databricks

Job description

Seeking an experienced Databricks Engineer to support the modernization of our enterprise data platform as part of our cloud transformation journey. The ideal candidate will design, develop, and optimize scalable data solutions on the Databricks Lakehouse Platform while enabling advanced analytics, AI/ML initiatives, and enterprise-wide data products.

This role will work closely with Data Architects, Data Scientists, Business Intelligence teams, and business stakeholders to deliver high-quality, governed, and reusable data assets supporting Operations, Transportation, Finance, Asset Management, Safety, and Corporate Services.

Key Responsibilities

Data Engineering & Development

  • Design, build, and maintain scalable data pipelines using Databricks, PySpark, and Spark SQL.
  • Develop ETL/ELT processes that ingest, transform, and curate large-scale structured and unstructured datasets.
  • Implement and support Medallion Architecture (Bronze, Silver, Gold layers).
  • Develop Delta Lake-based solutions with optimized performance and data quality controls.
  • Build batch and near real-time data processing solutions leveraging Spark Streaming and Kafka.
  • Create reusable frameworks and automation for data ingestion, monitoring, and orchestration.

Databricks Platform Management

  • Develop and maintain Databricks notebooks, workflows, jobs, and clusters.
  • Implement and manage Unity Catalog, access controls, and data governance standards.
  • Configure Delta Live Tables (DLT) and streaming pipelines.
  • Support environment promotion across Dev, QA, UAT, and Production.
  • Collaborate with platform teams on performance tuning and cost optimization initiatives.

Cloud & Integration

  • Integrate data from multiple enterprise sources including:

  • SAP
  • Teradata
  • Mainframe
  • DB2
  • SQL Server
  • Oracle
  • Tibco
  • Kafka
  • REST APIs

Design and implement cloud-native integrations supporting client data modernization strategy.

Work with Azure and AWS cloud services supporting Databricks workloads.

Data Quality & Governance

  • Implement data validation, reconciliation, and monitoring controls.
  • Develop automated data quality frameworks and exception handling.
  • Support data lineage, metadata management, and governance initiatives.
  • Ensure compliance with enterprise security and regulatory requirements.

Collaboration & Leadership

  • Engage with business teams to understand analytics requirements and translate them into scalable technical solutions.
  • Work closely with Data Scientists and BI developers to enable advanced analytics and reporting.
  • Mentor junior data engineers and establish engineering best practices.
  • Participate in Agile ceremonies and contribute to solution design discussions.

Requirements

Experience: 7+ years.

Technical Skills

· Databricks

  • Databricks Workspaces
  • Databricks Notebooks
  • Databricks Jobs & Workflows
  • Delta Lake
  • Delta Live Tables (DLT)
  • Unity Catalog
  • Databricks SQL
  • Databricks Asset Bundles (preferred)

Data Engineering

  • PySpark
  • Spark SQL
  • Python
  • SQL
  • ETL/ELT Development
  • Data Modeling
  • Data Warehousing Concepts

Streaming & Messaging

  • Apache Kafka
  • Spark Streaming
  • Event-driven Architectures

Cloud Platforms

  • Azure Databricks
  • Azure Data Factory
  • ADLS Gen2
  • Azure Key Vault
  • AWS S3
  • IAM
  • Cloud Data Services

Databases

  • Teradata
  • SQL Server
  • Oracle
  • DB2
  • BigQuery (preferred)
  • Snowflake (preferred)

DevOps

  • Git
  • CI/CD Pipelines
  • Azure DevOps
  • GitHub Actions
  • Infrastructure as Code (preferred)

Preferred Qualifications

  • Databricks Certified Data Engineer Associate or Professional.
  • Azure Data Engineer Associate certification.
  • Experience supporting Large Enterprise Data Modernization programs.
  • Experience migrating workloads from Teradata or legacy platforms to Databricks.
  • Familiarity with SAP Datasphere, SAP BW, or SAP BDC integrations.
  • Exposure to AI/ML workloads and feature engineering on Databricks.

Soft Skills

  • Strong analytical and problem-solving skills.
  • Ability to communicate technical concepts to business stakeholders.
  • Strong collaboration and teamwork skills.
  • Self-motivated with the ability to work independently.
  • Continuous learning mindset and passion for modern data technologies.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:05 min

Enhancing Databricks tooling for software engineering workflows

Alan Mazankiewicz · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

1:22 min

Analyzing differences between mobile and traditional backend DevOps

Mete Baydar Mete Baydar · World Congress 2025

2:38 min

Overview of Databricks and interactive data processing

Alan Mazankiewicz · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all