Mid-Level Data Engineer (On-Site in Washington, DC)

Agile5 Technologies, Inc.
Washington, DC, United States
23 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
9 years minimum
Compensation
$68,000.0 - $152,000.0
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Amazon Web Services Amazon S3 Apache HTTP Server Business Logic Microsoft Azure Code Review Databases Data Validation Information Engineering Extract Transform Load (ETL) Data Transformation
+22 more
Data Migration Data Profiling Apache Hadoop Apache Hive Python (Programming Language) Power BI Cloud Services SQL Databases SQL Server Integration Services Talend Esri GIS (Software) Informatica Powercenter Apache Spark Change Data Capture Git Data Lakes Pyspark Integration Tests Information Technology Software Version Control Data Pipelines Databricks

Job description

Description: The Mid-Level Data Engineer will support data migration, pipeline engineering, and modernization efforts for enterprise data lakehouse architectures. This role involves converting legacy Informatica artifacts into clean Python/PySpark code, migrating database schemas, and building automated data reconciliation pipelines. Working closely with senior engineering leadership and database managers, the ideal candidate will enforce high standards of data quality, data validation, and version control in a secure federal environment., * Execute daily data migration operations including data profiling, schema mapping, pipeline conversion, and automated reconciliation for Low and Medium complexity Informatica artifacts.

  • Convert Informatica mappings into well-documented Python/PySpark code, ensuring all business logic and data quality controls are preserved.
  • Migrate legacy Hive tables to Delta Lake format on S3 using Databricks ingestion tools.
  • Build and execute automated data reconciliation scripts to validate migration accuracy and establish Change Data Capture (CDC) pipelines for ongoing synchronization.
  • Commit all converted code into Azure DevOps with clear documentation and inline comments while maintaining Unity Catalog configurations.
  • Perform daily data profiling and side-by-side validation within legacy enclave environments.
  • Support Power BI and ESRI integration testing and validation.
  • Participate actively in peer code reviews, daily Agile ceremonies, and collaborative data validation sessions.
  • Contribute to Data Quality Assessment Reports and support training and knowledge transfer activities.
  • Performs other duties as assigned.

Security Clearance Requirements:

  • Public Trust / Tier 4 Eligible: No clearance required to apply; must be a U.S. citizen willing to undergo a background check to obtain a Public Trust / Tier 4 clearance.

Requirements

  • Minimum experience required varies by degree level: PhD with 0 years; Master’s degree with 3 years; Bachelor’s degree with 5 years; or High School Diploma with 9 years of relevant experience.
  • Proficiency in Python for data transformation and pipeline development, as well as SQL for query development and schema analysis.
  • Experience with ETL/ELT processes, data migration methodologies, and cloud data platforms (AWS, Azure, or GCP).
  • Familiarity with version control systems (Azure DevOps, Git) and data quality concepts including profiling, cleansing, and reconciliation.

Education Requirements: Bachelor’s degree in Computer Science, Data Engineering, Information Technology, or a related field is preferred (or equivalent combination of education and experience).

Desired Skills / Qualifications:

  • Experience with Databricks (notebooks, jobs, workspace navigation), PySpark, Apache Spark, and Delta Lake or Apache Iceberg table formats.
  • Proven track record converting visual ETL tools (Informatica, Talend, SSIS) to code-based pipelines.
  • Experience with Hive, HiveQL, or Hadoop ecosystem components.
  • Familiarity with federal IT environments, security requirements, and CI/CD pipelines for data engineering workflows.

About the company

About Agile5: Agile5 Technologies, Inc., is a Woman-Owned Small Business (WOSB) and Information Technology (IT) services firm that specializes in the design, development, testing, integration, and maintenance of enterprise software systems. We believe our employees are the company’s most valuable asset. We are invested in seeing our employees grow in their careers, while maintaining a work/life balance. We have an immediate, full-time need for a skilled, energetic, and driven Mid-Level Data Engineer.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all