> Markdown version of [/jobs/ext/2203362-databricks-data-engineer](https://www.wearedevelopers.com/jobs/ext/2203362-databricks-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Databricks Data Engineer - **Company:** Leidos, Inc. - **Location:** Decatur, GA, United States (Remote available) - **Experience:** Experienced - **Salary:** $87,100.0 - $157,450.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Airflow, Amazon Web Services, Microsoft Azure, Code Review, Continuous Integration, Data Architecture, Information Engineering, Data Governance, Data Integrity, Extract Transform Load (ETL), Data Transformation, Data Systems, Data Warehousing, DevOps, Dimensional Modeling, Distributed Computing Environment, Document-Oriented Databases, Apache Hive, Identity and Access Management, Python (Programming Language), Machine Learning, Metadata, Performance Tuning, Cloud Services, Data Streaming, Unstructured Data, Workflow Management Systems, Google Cloud, Data Ingestion, Apache Spark, Git, Containerization, Data Lakes, Pyspark, Information Technology, Software Coding, Stream Processing, Software Version Control, Data Pipelines, Databricks - **Published:** August 23, 2026 - **Apply:** https://www.careerjet.com/job/us2cebd1abbc015e43ba928dd2968a950b/eaa ## About the Role * Bachelor's degree in Computer Science, Information Technology, Engineering, Mathematics, or a related discipline. * 3-7+ years of experience in Data Engineering or ETL development. * Hands-on experience with Databricks and Apache Spark. * Strong proficiency in Python and PySpark. * Advanced SQL development skills. * Experience designing and developing enterprise-scale ETL/ELT pipelines. * Experience working with Delta Lake and Lakehouse architecture. * Strong understanding of distributed data processing concepts. * Experience with data modeling, dimensional modeling, and data warehousing. * Experience working with cloud platforms such as Microsoft Azure, AWS, or Google Cloud Platform. * Experience with Git-based source control and CI/CD deployment processes. * Strong analytical, troubleshooting, and problem-solving skills., * Databricks Certified Data Engineer Associate or Professional certification. * Experience with Unity Catalog, Delta Live Tables (DLT), and Databricks Workflows. * Experience with Structured Streaming and real-time data processing. ## Description Leidos is seeking a skilled Databricks Data Engineer to design, develop, and maintain scalable data engineering solutions on the Databricks Lakehouse Platform for the National Healthcare Safety Network (NHSN) support contract. This position is critical to supporting NHSN's modern data platform strategy by leveraging the Databricks Lakehouse Platform to build scalable, cloud-native data engineering solutions. The Databricks Data Engineer will enable efficient data ingestion, transformation, governance, and analytics while reducing operational complexity, improving data accessibility, and accelerating business insights. This role will help establish a modern enterprise data ecosystem capable of supporting advanced analytics, artificial intelligence, machine learning, and data-driven decision-making across the organization. The ideal candidate will have hands-on experience building enterprise-grade data pipelines using Apache Spark, PySpark, Delta Lake, and cloud-native technologies. This role will be responsible for developing high-performance ETL/ELT pipelines, implementing modern data architecture, optimizing distributed workloads, and enabling analytics and reporting across the organization. The successful candidate will work closely with data architects, business analysts, data scientists, and application teams to deliver reliable, secure, and scalable data solutions that support business intelligence, machine learning, and advanced analytics initiatives., * Design, develop, and maintain scalable ETL/ELT pipelines using Databricks and Apache Spark. * Build data ingestion frameworks to integrate structured, semi-structured, and unstructured data from various enterprise systems. * Develop data transformation logic using PySpark and Spark SQL. * Design, implement, and optimize Delta Lake tables for high performance, reliability, and scalability. * Implement Medallion Architecture (Bronze, Silver, Gold) for enterprise data processing. * Optimize Spark jobs through partitioning, caching, file compaction, and performance tuning techniques. * Develop and maintain data workflows using Databricks Workflows, Jobs, or orchestration tools such as Apache Airflow. * Implement data quality, validation, reconciliation, and monitoring processes to ensure data integrity. * Collaborate with cross-functional teams to understand business requirements and deliver high-quality data solutions. * Develop reusable data engineering frameworks, utilities, and coding standards. * Document data lineage, transformation logic, metadata, and technical designs. * Participate in code reviews and implement CI/CD processes using Git and DevOps tools. * Ensure compliance with enterprise data governance, security, and access management standards. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)