Senior Data Engineer

540
Arlington, VA, United States
12 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
9 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Airflow Amazon Web Services Data Analysis Automation of Tests Microsoft Azure Continuous Integration Information Engineering Data Infrastructure Extract Transform Load (ETL) Data Security Python (Programming Language)
+21 more
Metadata Meta-Data Management DataOps Software Engineering SQL Databases Data Streaming Google Cloud System Availability Apache Spark Data Lakes Pyspark Infrastructure Automation Frameworks Information Technology Apache Kafka Spark Streaming Data Management Machine Learning Operations Terraform Software Version Control Data Pipelines Databricks

Job description

540 is seeking a Senior Data Engineer to support a mission-critical technology modernization effort for the Department of War. You will lead the design and evolution of Databricks-based data pipelines and lakehouse capabilities that enable secure data integration, analytics, AI/ML, and operational workloads at enterprise scale.

Working with engineers, architects, cybersecurity teams, and mission stakeholders, you will translate complex requirements into secure, scalable solutions using Databricks, Apache Spark, and Delta Lake. You will define engineering standards, guide technical delivery, and mentor engineers while ensuring data quality, governance, and platform reliability., Lead the architecture and evolution of Databricks-based data pipelines and lakehouse capabilities Translate mission requirements into scalable data architectures and implementation strategies Define data engineering standards, reusable patterns, and best practices across engineering teams Architect automated ETL/ELT pipelines using Python, SQL, PySpark, Apache Spark, and Delta Lake Design scalable data models, schemas, data contracts, and medallion architecture patterns Lead the development of batch and streaming capabilities supporting operational, analytical, and AI/ML workloads Establish data quality, lineage, metadata, observability, and governance practices using Unity Catalog or similar technologies Optimize Databricks and Spark workloads for performance, scalability, reliability, and cost efficiency Establish CI/CD, infrastructure-as-code, testing, monitoring, and operational practices for Databricks environments Lead design reviews and resolve complex issues spanning data pipelines, infrastructure, and production services Partner with cybersecurity teams to incorporate security, access control, auditing, and governance requirements Communicate architecture decisions and mentor engineers on Databricks and data engineering best practices

Requirements

Citizenship & Clearance Requirement: Per client requirements, candidates must be U.S. Citizens with an active DoW Secret (or higher) clearance Education Requirement: Bachelor’s degree in Computer Science, Engineering, or a related technical field preferred; equivalent combinations of education and relevant experience will be considered 540 Internal Thrive Level: Senior Data Engineer, 9+ years of relevant data engineering or software engineering experience Experience leading enterprise-scale data platform and pipeline implementations using Databricks Advanced proficiency with Python, SQL, PySpark, Apache Spark, and Delta Lake Experience architecting large-scale ETL/ELT, batch, and streaming pipelines Experience designing lakehouse architectures, data models, schemas, and data contracts Experience managing and optimizing Databricks jobs, workflows, compute resources, and Spark workloads Experience implementing data quality, monitoring, lineage, metadata management, and governance capabilities Experience with Unity Catalog or similar data-governance and access-control solutions Experience operating Databricks within AWS, Azure, or Google Cloud Experience with CI/CD, infrastructure as code, automated testing, and source control Strong understanding of data security, privacy, governance, and access-control principles Experience leading technical reviews, mentoring engineers, and communicating architecture decisions

NICE TO HAVE

Relevant Databricks certification Experience supporting DoW, federal, Advana, or other mission data environments Experience building data platforms in classified, regulated, or mission-critical environments Experience with Terraform, Databricks Asset Bundles, Airflow, or similar automation and orchestration tools Experience architecting streaming solutions using Kafka, Kinesis, Pulsar, or Spark Structured Streaming Experience supporting AI/ML pipelines, feature platforms, or MLflow Currently holds, or is willing to obtain within 30 days of employment, an approved certification such as Cloud+, GSEC, Security+, or SSCP

Benefits & conditions

Flexible PTO + all Federal holidays off Health, dental and vision insurance plans Flexible Spending Account (FSA) 401k with employer match Company-sponsored life insurance, short- and long-term disability Professional development (training, certifications, conferences) Paid cloud developer accounts Referral bonuses HQ office perks (parking / metro reimbursement, nitro coffee & lunches) Annual social events (540 Week, hackathon, charity golf tournament, etc.) Access to 540’s Washington Capitals & Nationals tickets

EQUAL EMPLOYMENT OPPORTUNITY (EEO)

540’s policy is to provide equal employment opportunity to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.

This policy applies to all terms and conditions of employment, including recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation and training.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.clearancejobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:00 min

Separating dataset creation from low-level software implementation steps

Jan Zawadzki · WWC 2022

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all