Databricks Architect

Innovative Information Technologies, Inc
United States
4 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Temporary to permanent
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Query Performance Amazon Web Services Business Logic Automation of Tests Microsoft Azure Big Data Code Review Continuous Integration Data Architecture Information Engineering Software Debugging DevOps
+22 more
Identity and Access Management Python (Programming Language) Performance Tuning Azure Active Directory Data Mesh Standard Sql SQL Databases Google Cloud Multi-Cloud Cloudformation Data Lakes Pyspark Data Lineage Data Management Machine Learning Operations Software Coding Terraform Domain Driven Design Software Version Control Data Pipelines Azure Resource Manager Databricks

Requirements

Current, hands-on development experience is essential. This role requires the ability to read, write, debug, and lead code reviews of production PySpark and SQL code not just architectural oversight. Candidates must demonstrate recent (within past 12 months) hands-on development work including debugging live code, explaining business logic and technical implementation, optimizing queries, and implementing data pipelines. Ability to comprehend and explain unfamiliar code samples is a core requirement. Advanced proficiency in SQL and Python/PySpark with demonstrated ability to write complex transformations, optimize query performance, explain code logic at both business and technical levels, and troubleshoot production issues. Must be comfortable with CTEs, window functions, table-valued functions, APPLY operators, advanced SQL operators, and PySpark DataFrame/RDD operations. Current, active coding skills required not aspirational or theoretical knowledge. 7+ years of experience in data architecture, data engineering, or platform engineering roles, with at least 3+ years focused on Databricks platform architecture. Expert-level knowledge of Databricks platform components: Unity Catalog, Delta Lake, Delta Live Tables, Workflows, SQL Warehouses, MLflow, and Databricks SQL. Deep expertise in Unity Catalog governance, including metastore design, catalog/schema strategies, permission models, data lineage, and multi-workspace/multi-cloud patterns. Strong architectural background in cloud platforms (Azure, AWS, or Google Cloud Platform), including storage services, identity management (Azure AD, AWS IAM), networking, and security best practices. Proven experience designing enterprise-scale data architectures, including medallion/multi-hop architectures, data mesh patterns, domain-driven design, and data product frameworks. Hands-on experience with infrastructure-as-code (Terraform, ARM templates, CloudFormation) for platform configuration and governance automation. Strong understanding of DevOps practices, CI/CD pipelines, version control strategies, and automated testing for data platforms. Experience with performance tuning, cost optimization, and capacity planning for large-scale data platforms.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:05 min

Enhancing Databricks tooling for software engineering workflows

Alan Mazankiewicz · LIVE

1:22 min

Analyzing differences between mobile and traditional backend DevOps

Mete Baydar Mete Baydar · World Congress 2025

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

Videos

See all

Related articles

See all