Databricks Solution Architect
Voto Consulting LLC
New York, NY, United States
2 months ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source
Tech stack
Agile Methodology
Microsoft Azure
Big Data
Software Quality
Continuous Integration
Data Architecture
Information Engineering
Data Governance
Extract Transform Load (ETL)
Data Warehousing
Software Design Patterns
Python (Programming Language)
+20 more
Key Management
Performance Tuning
Scrum Methodology
Role-Based Access Control
Release Management
Azure Data Lake
SQL Databases
Enterprise Data Management
Cloud Platform System
Azure Data Factory
Apache Spark
Caching
Data Lakes
Pyspark
Information Technology
Collibra
Data Management
Data Pipelines
Serverless Computing
Databricks
Job description
- Lead the end-to-end architecture and solution design for enterprise data platforms on Azure, with strong focus on Databricks Lakehouse, Delta Lake, and scalable cloud-native data ecosystems.
- Define target-state data architecture, ingestion patterns, transformation frameworks, and serving layers to support reporting, advanced analytics, ML, and business-critical decisioning use cases.
- Design and implement robust ETL/ELT pipelines using PySpark, SQL, Databricks Workflows, Auto Loader, and Delta Live Tables for batch and near real-time processing.
- Own architecture standards for data modeling, medallion design, reusable engineering patterns, CI/CD, code quality, environment strategy, and release management across Databricks solutions.
- Drive platform governance and security using Unity Catalog, RBAC/ABAC controls, lineage, auditability, and integration with enterprise governance services such as Purview.
- Optimize solution performance by tuning Spark workloads, cluster policies, partitioning strategy, file sizing, caching, and compute cost management for large-scale data processing.
- Collaborate with business stakeholders, product owners, analysts, architects, and downstream consumers to translate functional and non-functional requirements into scalable technical designs.
- Provide technical leadership to engineering teams by reviewing designs, guiding implementation, resolving architectural bottlenecks, and establishing best practices for Databricks-based delivery.
- Evaluate and recommend Databricks capabilities such as Photon, serverless compute, Lakehouse Federation, and streaming patterns to improve scalability, maintainability, and time to value.
- Ensure strong delivery governance through estimation, technical planning, dependency management, risk mitigation, and Agile execution including sprint planning, backlog refinement, and design reviews.
Requirements
- 10-15 years of experience in data engineering, cloud data platform design, or enterprise data architecture, with at least 5+ years of strong hands-on experience on Databricks and Azure.
- Bachelor’s degree in computer science, Information Technology, Engineering, or a related discipline; master’s degree is preferred.
- Strong expertise in designing modern data architectures, including Lakehouse, medallion architecture, data modeling, data warehousing, and scalable ingestion and transformation frameworks.
- Deep technical proficiency in Databricks, PySpark, Python, SQL, Delta Lake, Databricks Workflows, Auto Loader, and Delta Live Tables.
- Strong experience with Azure cloud services such as Azure Data Factory, Azure Data Lake Storage, Azure Key Vault, Azure DevOps, and integration of Databricks with broader enterprise cloud ecosystems.
- Hands-on experience in defining architecture standards, reusable design patterns, CI/CD strategy, environment management, and delivery best practices for enterprise-scale data platforms.
- Experience with data governance, lineage, and security frameworks including Unity Catalog, role-based access controls, and integration with enterprise governance tools such as Azure Purview.
- Proven ability to optimize large-scale Spark and Databricks workloads, including performance tuning, cluster sizing, workload management, and cost optimization.
- Experience engaging with business stakeholders, product owners, architects, and engineering teams to convert business requirements into scalable solution designs and implementation roadmaps.
- Strong communication, leadership, and problem-solving skills with the ability to mentor teams, review technical designs, and drive architecture decisions across cross-functional programs.
- Preferred: Knowledge of insurance domain data models, regulatory considerations, and analytics use cases relevant to underwriting, claims, pricing, or risk functions.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on dice.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
BB
Benedikt Bischof
about 4 years ago
DS
Dhannush Subramani
Top Big Data Technologies That You Need to Know
about 4 years ago
CH
Chris Heilmann
Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production
almost 2 years ago
MH
Michael Hunger
Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?
7 months ago
AJ
Austin Joy
What Are The Top Skills Required For Azure Developers?
over 4 years ago
EM
Eli McGarvie
Data Engineer Salary UK
about 3 years ago