> Markdown version of [/jobs/ext/2026796-databricks-data-engineer](https://www.wearedevelopers.com/jobs/ext/2026796-databricks-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Databricks Data Engineer - **Company:** Peraton Inc - **Location:** United States (Remote available) - **Experience:** Experienced - **Salary:** $112,000.0 - $179,000.0 - **Contract:** Permanent contract - **Skills:** Airflow, Amazon Web Services, Microsoft Azure, Big Data, Business Systems, Cloud Computing, Continuous Integration, Data Dictionary, Data Governance, Data Structures, Apache Hive, Python (Programming Language), Meta-Data Management, Enterprise Software Applications, Data Ingestion, Git, Data Lakes, Pyspark, Infrastructure Automation Frameworks, Data Lineage, Optimization Algorithms, Enterprise Integration, Machine Learning Operations, Tools for Reporting, Data Pipelines, Legacy Systems, Databricks - **Published:** August 11, 2026 - **Apply:** https://www.clearancejobs.com/jobs/9085680/databricks-data-engineer ## About the Role * 5 years work experience with BS/BA; 3 years with MS/MA * US Citizenship * Active DoD Secret clearance * 2 years working on the Databricks platform * Strong proficiency in PySpark, Spark SQL, and Python for large-scale data processing and pipeline development * Hands-on experience with Unity Catalog administration, including metastore management, access policies, and data lineage * Experience with pipeline orchestration tools (Airflow, Databricks Workflows) * Experience with Delta Lake, schema evolution, time travel, and optimization techniques * Experience integrating heterogeneous enterprise systems, including legacy/custom integrations * Familiarity with cloud platforms (Azure/AWS) and infrastructure-as-code practices * Understanding of data governance principles, compliance frameworks, and CUI/PII handling practices * Familiarity with financial, HR, CRM, or ITSM data structures * Git/CI-CD pipeline experience * DoD 8570 certification, * Databricks Certified Data Engineer Associate or Professional certification ## Description The Databricks Data Engineer will be responsible for the hands-on build-out of a Government-owned Databricks workspace and the data ingestion/integration work needed to consolidate agency business system data. This includes designing and implementing data pipelines from core enterprise systems, implementing Unity Catalog for data governance, configuring MLflow for machine learning workflows, and developing comprehensive documentation to support sustainment beyond the pilot. This role is 100% Remote. Other Responsibilities Include: * Build out and configure a dedicated Government workspace within the existing Databricks environment * Design and implement data ingestion pipelines from core agency business systems including financial, HR, CRM, and ITSM systems * Leverage native/built-in connectors where source systems support them; design custom integration approaches for legacy systems * Normalize and prepare ingested data within Databricks for consumption by downstream visualization/reporting tools * Implement Databricks Unity Catalog for centralized data governance, metadata management, active auditing, and end-to-end lineage tracking * Review current configuration, assess security controls for CUI/PII/PHI/financial data, and implement improvements * Develop comprehensive "as-built" documentation including physical/logical architecture diagrams, automated data dictionaries, and SOPs * Document data sources, integration methods, and data lake architecture decisions to support sustainment ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Dev Digest 162: AI careers, MCP, AWS best practices & floppy sweaters](https://www.wearedevelopers.com/magazine/571-dev-digest-162-ai-careers-mcp-aws-best-practices-floppy-sweaters)