> Markdown version of [/jobs/ext/1935661-pyspark-data-engineer](https://www.wearedevelopers.com/jobs/ext/1935661-pyspark-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # PySpark Data Engineer - **Company:** Kforce Inc. - **Location:** Arlington, DC, United States - **Salary:** $124,800.0 - $176,800.0 - **Contract:** Temporary contract - **Skills:** Data Analysis, Automation of Tests, Big Data, Data Architecture, Data Governance, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Systems, DevOps, Distributed Computing Environment, Performance Tuning, Cloud Services, Software Deployment, SQL Databases, Enterprise Data Management, Data Processing, Git, SC Clearance, Containerization, Pyspark, Qlikview, Data Management, Data Pipelines, Docker, Databricks - **Published:** August 5, 2026 - **Apply:** https://www.clearancejobs.com/jobs/9075999/pyspark-data-engineer ## About the Role Active TS/SCI clearance preferred; candidates with an active Secret clearance will also be considered. Strong experience working within Databricks environments. Advanced proficiency with: PySpark SQL ETL development and optimization Data pipeline design and engineering Experience implementing CI/CD pipelines and DevOps best practices. Hands-on experience with Git version control. Experience deploying and managing containerized applications using Docker and Kubernetes. Strong understanding of distributed data processing and enterprise-scale data architectures. Ability to optimize and tune high-volume data processing workloads. Excellent problem-solving and analytical skills., Experience supporting federal or defense programs. Databricks Certified Data Engineer Associate or Professional certification. Experience working with financial management data. Qlik dashboard and reporting experience. Experience with secure federal data platforms such as Advana or Warfighter Data Platform (WDP). Familiarity with cloud-native data engineering architectures and modern data integration frameworks. Technical Environment Databricks PySpark SQL Docker Kubernetes Git CI/CD Pipelines DevOps Enterprise Data Platforms Qlik Analytics ## Description We are seeking a Data Engineer to support a high-impact defense program focused on delivering advanced data solutions for mission-critical operations. This position offers the opportunity to work with modern cloud and big data technologies, building scalable data architectures and pipelines that support enterprise analytics, data integration, and decision-making across a secure environment. The ideal candidate will bring strong expertise in Databricks, PySpark, SQL, and DevOps practices, along with experience designing and optimizing large-scale data processing solutions., Architect, develop, and maintain scalable data pipelines supporting enterprise data ingestion, processing, and transformation. Design and optimize ETL workflows for large, complex datasets using PySpark and SQL. Build and manage data engineering solutions within the Databricks platform. Implement performance tuning strategies to improve reliability, scalability, and efficiency of data processing workloads. Develop and maintain CI/CD pipelines to support automated testing, deployment, and monitoring. Containerize and deploy applications using modern technologies such as Docker and Kubernetes. Collaborate with software engineers, analysts, architects, and stakeholders to deliver scalable data solutions. Apply DevOps principles and best practices to streamline development and deployment processes. Support data governance, data quality, and system performance initiatives. Troubleshoot and resolve data pipeline, platform, and performance-related issues. ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)