> Markdown version of [/jobs/ext/2296941-data-engineer-databricks](https://www.wearedevelopers.com/jobs/ext/2296941-data-engineer-databricks). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer, Databricks - **Company:** KeyLogic, LLC - **Location:** Washington, DC, United States - **Experience:** Experienced - **Salary:** $90,000.0 - $140,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Data Analysis, Big Data, Cloud Computing, Continuous Integration, Data Integration, Data Systems, Python (Programming Language), Performance Tuning, Data Logging, Data Processing, Apache Spark, Data Lakes, Infrastructure Automation Frameworks, Terraform, Stream Processing, Software Version Control, Data Pipelines, Databricks - **Published:** August 29, 2026 - **Apply:** https://diversityjobs.com/main/sendform/8/8/28176/1/18123673?backUrl=%2Fcareer%2F18123673%2FData-Engineer-Databricks-D-C-Washington ## About the Role * U.S. Citizenship * 2+ years of engineering experience on Databricks and large-scale big data projects (Delta Lake, DLT, Apache Spark, unity Catalog). * Bachelor's Degree * 4+ years of experience in Python, Java, or equivalent. Comfortable with evaluation tooling, logging, monitoring, and observability * Experience with government contracting ## Description As a Data Engineer in this role, you will be responsible for designing, building, and maintaining robust data pipelines that support the organization's analytical and operational needs. You will work closely with cross-functional teams to ensure data is collected, processed, and made available in a reliable and scalable manner. Leveraging your expertise, you will optimize data workflows and implement best practices for data integration, transformation, and storage. Your contributions will directly impact the efficiency and effectiveness of data-driven decision-making across the business. To excel in this position, you will utilize tools such as Databricks for advanced data processing and analytics, and Terraform for infrastructure automation and management. Familiarity with cloud-based environments, version control systems, and CI pipelines will further enhance your ability to deliver high-quality solutions. Experience with large-scale data architectures, real-time data streaming, and performance tuning will help you address complex challenges and drive continuous improvement. Your background in collaborating with diverse technical teams and managing end-to-end data solutions will be instrumental in achieving success in this dynamic environment. Position Responsibilities: * Design, build, and maintain scalable data pipelines using Databricks * Develop and implement infrastructure as code solutions with Terraform * Collaborate with cross-functional teams to understand data requirements and deliver robust solutions * Optimize data workflows for performance, reliability, and scalability * Monitor, troubleshoot, and resolve issues related to data processing and infrastructure * Ensure data quality and integrity throughout the data lifecycle * Automate repetitive tasks to improve efficiency and reduce manual intervention ## Related Videos - [OLTP in the Lakehouse: Redefining Data for AI Workloads](https://www.wearedevelopers.com/videos/2038-oltp-in-the-lakehouse-redefining-data-for-ai-workloads) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Cutting LLM Costs Without Cutting Quality: How to Beat Proprietary LLMs with Fine-Tuned Open Source](https://www.wearedevelopers.com/videos/100151-cutting-llm-costs-without-cutting-quality-how-to-beat-proprietary-llms-with-fine-tuned-open-source) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering)