Databricks Data Engineer

Radancy
Washington, DC, United States
12 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$113,000.0 - $188,000.0
Working hours
Regular working hours
Job source

Tech stack

Unity 3d Artificial Intelligence Amazon Web Services Microsoft Azure Big Data Cloud Computing Cloud Storage Code Review Continuous Integration Data Architecture Data Validation Information Engineering
+26 more
Data Governance Data Infrastructure Extract Transform Load (ETL) Data Transformation Data Warehousing DevOps Digital Assets Python (Programming Language) Performance Tuning Software Tools Cloud Services Standard Sql SQL Databases Data Streaming Workflow Management Systems Data Processing Cloud Platform System Apache Spark Containerization Data Lakes Pyspark Information Technology Data Delivery Terraform Data Pipelines Databricks

Job description

Guidehouse is seeking a Databricks Data Engineer to join our AI & Data team to support client projects involving large-scale data pipelines, data transformation, platform modernization, and advanced analytics and AI solution delivery. This role focuses on designing, building, and maintaining scalable data solutions using Databricks, Spark, SQL, Python, Delta Lake, and related cloud-native technologies. The role requires hands-on technical expertise, strong problem-solving skills, and the ability to collaborate with clients and cross-functional teams to deliver reliable, secure, and well-documented data engineering solutions., * Design, build, and maintain scalable data pipelines using Databricks, PySpark, SQL, Delta Lake, and related cloud-native data engineering tools.

  • Develop and support batch and streaming data workflows, including ingestion, transformation, validation, and publishing of curated data products.
  • Write, optimize, and maintain Python, PySpark, and SQL code for data processing, orchestration, and performance tuning.
  • Work with large-scale datasets using Databricks, Spark, Delta Lake, Unity Catalog, and cloud storage services.
  • Troubleshoot and resolve data pipeline failures, performance issues, data quality issues, and workflow bottlenecks.
  • Translate business and technical requirements into data engineering designs, processing scripts, and job orchestration workflows.
  • Perform data validation, quality checks, code reviews, and issue resolution to support accurate and reliable data products.
  • Collaborate with cross-functional teams including solution architects, data scientists, analysts, DevOps engineers, and client stakeholders.
  • Communicate technical concepts, delivery impacts, and data engineering considerations to technical and non-technical audiences.
  • Document pipelines, datasets, data models, workflows, and engineering decisions to support maintainability, transparency, and reuse.
  • Follow data governance, security, lineage, and compliance standards within the Databricks platform.

Requirements

  • Bachelor’s degree in computer science, engineering, mathematics, statistics, or another relevant field.
  • 3-8 years of relevant experience in data engineering, data architecture, or cloud data platform implementation.
  • Strong experience with Python, PySpark, and SQL for data transformation, pipeline development, and data processing.

  • Experience with Databricks, Spark, Delta Lake, Unity Catalog, or similar cloud-native data platforms.
  • Experience developing batch or streaming data pipelines, ETL/ELT workflows, and reusable data assets.
  • Experience with data modeling, data warehousing, data quality validation, and large-scale data processing concepts.
  • Ability to troubleshoot technical issues, communicate engineering recommendations clearly, and work effectively in team-based delivery environments.
  • Experience applying data governance, access control, lineage, and security practices within cloud data platforms.

What Would Be Nice to Have:

  • 2+ years of hands-on experience with the Databricks platform.
  • Experience using Unity Catalog for data governance, schema design, data modeling standards, access controls, lineage, and secure management of enterprise data assets.
  • Experience designing and building standard medallion data architectures for scalable ingestion, transformation, quality validation, and curated data delivery.
  • Active Databricks Data Engineer Associate, Databricks Data Engineer Professional, or related certification.
  • Experience with CI/CD tools and practices for notebooks, workflows, orchestration jobs, and infrastructure deployment.
  • Experience with cloud platforms such as Azure, AWS, or GCP.
  • Experience working in project-based or consulting delivery environments.
  • Familiarity with streaming frameworks, orchestration tools, Terraform, or related data platform administration capabilities.

Benefits & conditions

The annual salary range for this position is $113,000.00-$188,000.00. Compensation decisions depend on a wide range of factors, including but not limited to skill sets, experience and training, security clearances, licensure and certifications, and other business and organizational needs.

What We Offer:

Guidehouse offers a comprehensive, total rewards package that includes competitive compensation and a flexible benefits package that reflects our commitment to creating a diverse and supportive workplace.

About Guidehouse

Guidehouse is an Equal Opportunity Employer-Protected Veterans, Individuals with Disabilities or any other basis protected by law, ordinance, or regulation.

Guidehouse will consider for employment qualified applicants with criminal histories in a manner consistent with the requirements of applicable law or ordinance including the Fair Chance Ordinance of Los Angeles and San Francisco.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

2:36 min

Development tools for spatial computing and drones

Zaid Zaim Zaid Zaim · WWC 2023

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all