Data Engineer - Junior, Mid and Senior Secret Clearance

DUNHILL PROFESSIONAL SEARCH
Washington, DC, United States
12 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Starter
Experience required
2 years minimum
Compensation
$70,000.0 - $130,000.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Data Analysis Microsoft Azure BigQuery Cloud Engineering Profiling Continuous Integration Data Architecture Data Validation Information Engineering Data Governance Data Security
+27 more
Data Warehousing Digital Assets Distributed Computing Environment Distributed Systems Document-Oriented Databases Meta-Data Management Data Streaming Unstructured Data Web Application Frameworks Google Cloud Azure Data Factory System Availability Snowflake Apache Spark Git SC Clearance Containerization Data Lakes Git Flow Information Technology Data Lineage AWS Glue Data Analytics Apache Kafka Terraform Data Pipelines Serverless Computing

Job description

The Data Engineer Journeyman designs, builds, and operates scalable data pipelines and platforms that ingest, process, and store structured and unstructured data to support mission-critical use across the enterprise data environment. Leveraging modern batch and streaming frameworks, this role develops optimized data models, transformation logic, and storage solutions that enable analytics, reporting, and advanced data use cases for business and technical stakeholders. The Data Engineer Journeyman also implements data quality, lineage, and governance practices while ensuring security and compliance in a highly regulated federal data context.

This position collaborates with cross-functional teams-including data scientists, analysts, and security stakeholders-to understand data requirements, refine data workflows, and continuously improve platform reliability, performance, and resilience. The engineer troubleshoots pipeline issues, documents data architecture and pipelines, and contributes to ongoing modernization and automation of data engineering processes and tooling across the client environment., * Design, develop, and maintain batch and streaming data pipelines using frameworks such as Apache Spark, Kafka, or equivalent cloud-native services to support high-volume ingestion and processing for mission-critical workloads.

  • Build and optimize data models and schemas (including star, snowflake, and normalized designs) that support analytical, reporting, and operational use cases across the enterprise.
  • Implement data validation, profiling, and monitoring capabilities to ensure high data quality, integrity, and reliability across all stages of the data pipeline lifecycle.
  • Design, deploy, and tune scalable storage platforms such as data lakes and data warehouses, employing partitioning, indexing, and compression strategies to improve performance and cost efficiency.
  • Establish and maintain metadata management and data lineage tracking using catalog and governance tools to provide transparency, auditability, and regulatory compliance for data assets.
  • Apply secure data engineering practices, including encryption, role-based access controls, and adherence to government data standards and policies in a highly regulated environment.
  • Automate CI/CD workflows for data pipelines and related infrastructure using tools such as Git, Terraform, and containerization technologies to enable repeatable, reliable deployments.
  • Troubleshoot and resolve pipeline failures, latency issues, and performance bottlenecks in distributed computing environments, driving root cause analysis and long-term remediation.
  • Collaborate with data scientists, analysts, and business stakeholders to understand data requirements, refine transformation logic, and ensure that data products meet analytical and operational needs.
  • Document data architectures, pipeline designs, and operational procedures, and contribute to data governance activities and continuous improvement of data engineering standards and practices.

Requirements

  • Bachelor’s degree in Computer Science, Information Technology, Data Engineering, or a closely related field, or equivalent relevant experience.
  • Typically 2-8 years of professional experience in data engineering or a closely related field, including hands-on work with data pipelines, data models, and distributed data processing.
  • Demonstrated experience building and operating batch and/or streaming data pipelines using modern frameworks (e.g., Apache Spark, Kafka) or cloud-native equivalents.
  • Experience designing and implementing relational and analytical data models, including star, snowflake, and normalized schemas, in support of reporting and analytics.
  • Practical experience implementing data quality controls, validation, and monitoring, as well as using metadata and catalog tools to manage data assets and lineage.
  • Demonstrated ability to apply secure data engineering practices, including encryption, access control, and compliance with government or enterprise security standards.
  • Active Secret Security Clearance Required
  • U.S. citizenship required.
  • Willingness and ability to work effectively in a remote, distributed team environment supporting federal or enterprise IT operations.
  • Candidate must be within 50 miles of Washington, DC, * Experience with one or more major cloud platforms (AWS, Azure, or Google Cloud) and their native data engineering services (e.g., AWS Glue, Azure Data Factory, BigQuery, or similar).
  • Familiarity with data governance frameworks and regulatory compliance requirements applicable to federal or highly regulated environments.
  • Relevant data engineering or cloud certifications (e.g., AWS Certified Data Analytics, Azure Data Engineer Associate, or equivalent).
  • Experience automating CI/CD for data pipelines and infrastructure using Terraform, Git-based workflows, and containerization/orchestration technologies.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.clearancejobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:02 min

Mapping Git flow branches to application tester segments

Majid Hajian · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:30 min

Leveraging BigQuery ML for scalable SQL-based segmentation experiments

Julian Joseph · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

5:19 min

Audience Q&A on future tools and Git commands

PJ Hagerty PJ Hagerty · WWC Europe 2026

3:53 min

Introduction to git flow and clean feature branches

Johannes Haux · WWC 2022

Videos

See all

Related articles

See all