Senior Data Engineer - Databricks

SugarCRM
Denver, CO, United States
4 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours

Tech stack

Amazon Web Services Microsoft Azure Continuous Integration Customer Data Management Information Engineering Data Governance Extract Transform Load (ETL) Data Security Python (Programming Language) PostgreSQL Performance Tuning Role-Based Access Control
+11 more
SQL Databases SQL Server Integration Services Data Streaming Apache Spark Build Management Data Lakes Pyspark Data Delivery Software Version Control Data Pipelines Databricks

Job description

Experteer Overview In this Senior Data Engineer role, you will own and optimize the Databricks-based data pipelines that power Sugar Predict, ensuring reliable, scalable data delivery across Azure and AWS. You’ll collaborate with ML, product, and Enterprise Architecture teams to maintain a fast, clean data backbone that supports global scale. You’ll reduce compute costs, modernize legacy pipelines, and enforce security and data quality across multi-tenant environments. This position offers hands-on impact on how revenue intelligence is delivered to mid-market customers and involves on-call incident response and cross-functional work to drive platform growth. Compensation / Benefits * Own Databricks production support across all data flows, including monitoring, alerting, and incident response * Maintain and report SLA metrics for data pipeline delivery and platform health * Optimize pipelines to reduce Databricks compute costs and improve throughput * Migrate legacy ETL/ELT pipelines to Databricks with automation to minimize manual intervention * Onboard and harden tenant data pipelines for new customers to ensure isolated, reliable data from day one * Design and build high-performance Databricks pipelines ingesting and transforming ERP and CRM data at scale across Azure and AWS * Own Delta Lake architecture including schema design, partitioning, quality enforcement, and incremental processing * Enforce data security best practices across Databricks (RBAC, secrets, Unity Catalog, Attestation Maps, compliance) * Implement data quality monitoring and observability to support model inputs and prediction accuracy * Apply multi-tenant isolation patterns for reliable, secure data delivery across enterprise customers * Partner with Enterprise Architecture to integrate pipelines with the SugarAI product ecosystem * Support globally distributed operations with on-call rotation and after-hours incident response * Maintain documentation, runbooks, and architectural decision records for operational readiness * Apply CI/CD practices to data pipelines with version control, testing, and deployment tooling Tasks * 4+ years of data engineering experience * 2+ years on Databricks or Apache Spark in Azure and/or AWS * Proficiency in PySpark, SQL, and Python with production-grade pipelines under SLA * Hands-on experience with Delta Lake (schema evolution, ACID, optimize/vacuum, incremental/streaming) * Pipeline performance tuning and compute optimization in Databricks * PostgreSQL knowledge for production pipelines * Experience maintaining legacy ETL tooling (SSIS, Informatica, custom pipelines) * Experience with large-scale multi-tenant architectures emphasizing isolation and data privacy * Ability to collaborate across data science, product, and infrastructure teams * Strong understanding of data governance, security, and compliance across multi-tenant environments Key requirements * Excellent healthcare package for you and your family * 401(k) match * Unlimited Paid Time Off * Paid Parental Leave * Online Legal Services (Rocket Lawyer) * Financial Planning Services (Origin)

Requirements

for operational readiness * Apply CI/CD practices to data pipelines with version control, testing, and deployment tooling Tasks * 4+ years of data engineering experience * 2+ years on Databricks or Apache Spark in Azure and/or AWS * Proficiency in PySpark, SQL, and Python with production-grade pipelines under SLA * Hands-on experience with Delta Lake (schema evolution, ACID, optimize/vacuum, incremental/streaming) * Pipeline performance tuning and compute optimization in Databricks * PostgreSQL knowledge for production pipelines * Experience maintaining legacy ETL tooling (SSIS, Informatica, custom pipelines) * Experience with large-scale multi-tenant architectures emphasizing isolation and data privacy * Ability to collaborate across data science, product, and infrastructure teams * Strong understanding of data governance, security, and compliance across multi-tenant environments Key requirements * Excellent healthcare package for you and your family * 401(k) match * Unlimited Paid a and Off * Paid Parental Leave * Online Legal Services (Rocket Lawyer) * Financial Planning Services (Origin)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:24 min

The governance failures of centralized data lakes

Mario Meir-Huber · LIVE

2:27 min

Managing traffic and tracking costs with Databricks Unity Catalog

Viktoria Semaan Viktoria Semaan · WWC Europe 2026

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:59 min

Key takeaways and accessing the Databricks developer toolkit

Viktoria Semaan Viktoria Semaan · WWC Europe 2026

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all