Data Engineer

Marathon TS Inc
Washington, DC, United States
5 days ago
Apply on www.clearancejobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Agile Methodology Airflow Amazon Web Services Amazon S3 Data Analysis Computing Platforms Big Data Collaborative Software Information Systems Databases Data Architecture
+42 more
Data Dictionary Information Engineering Data Governance Data Infrastructure Data Integration Extract Transform Load (ETL) Data Stores Data Systems DevOps Distributed Computing Environment Amazon DynamoDB Elasticsearch Push Technology Java Message Service (JMS) Python (Programming Language) PostgreSQL MongoDB Object-Oriented Software Development Open Source Technology Raw Data Redis Software Tools Amazon Simple Notification Service (SNS) SQL Databases Data Streaming Systems Architecture Scripting Cloud Platform System AWS Lambda Pandas Kubernetes Atlassian Tools Apache Kafka Apache Nifi Data Management 3-tier Architectures Data Delivery Amazon Simple Queue Service (SQS) Data Pipelines Docker Server Operating Systems & Platforms Amazon Redshift

Job description

As a Data Engineer, you will be required to interpret business needs and select appropriate technologies and have experience in implementing data governance of shared and/or master sets of data. You will work with key business stakeholders, IT experts, and subject-matter experts to plan and deliver optimal data solutions. You will create, maintain, and optimize data pipelines as workloads move from development to production for specific use cases to ensure seamless data flow for the use case. You will perform technical and non-technical analyses on project issues and help to ensure technical implementations follow quality assurance metrics. You will analyze data and systems architecture, create designs, and implement information systems solutions., * Identify, design and implement internal process improvements including re-designing infrastructure for greater scalability, optimizing data delivery, and automating manual processes.

  • Develop and design data pipelines to support an end-to-end solution.
  • Develop and maintain artifacts (e.g., schemas, data dictionaries, and transforms related to ETL processes).
  • Integrate data pipelines with AWS cloud services to extract meaningful insights.
  • Work with stakeholders to support their data infrastructure needs and assist with data-related technical issues.
  • Design and develop robust and functional dataflows to support raw data and expected data.
  • Provide Tier 3 technical support for deployed applications and dataflows.
  • Define and communicate a clear product vision for our client’s software products, aligning user needs and business objectives.
  • Create and manage product roadmaps that reflect both innovation and growth strategies.
  • Partner with a government product owner and a product team of 7-8 FTEs.
  • Collaborate with the rest of data engineering team to design and launch new features.
  • Coordinate and document dataflows, capabilities, etc.
  • Occasionally (as needed) support to off-hours deployment such as evening or weekends.

Requirements

  • Bachelor’s degree or equivalent practical experience.
  • Expertise in distributed computing frameworks to handle large-scale data processing.
  • Familiarity working with Amazon Web Managed Services (AWS) or any other cloud environments.
  • Working experience with datastores like PostgreSQL, S3, Redshift, MongoDB/DynamoDB, Redis, Elasticsearch/OpenSearch and SQL.
  • Proficient utilizing Python with key libraries like pandas and PySark, NiFi, Airflow, AWS Lambda or similar technologies.
  • Working knowledge with software platforms and services, such as Docker, Kubernetes, JMS/SQS, SNS and Kafka.
  • Familiar with Linux/Unix server environments.
  • Experience with Agile development methodology.
  • Publishing and/or presenting design reports.
  • Coordinating with other team members to reach project milestones and deadlines.
  • Working knowledge with Collaboration tools, such as, Jira and Confluence., * Master’s degree or equivalent experience in a related field.
  • Familiarity and experience with the Intelligence Community (IC), and the Client cycle.
  • Familiarity and experience with the Department of Homeland Security (Client).
  • Direct Experience with Client and Intelligence Community (IC) component’s data architectures and environments (IC-GovCloud experience preferred).
  • Experience with cloud message APIs and usage of push notifications.
  • Keen interest in learning and using the latest software tools, methods, and technologies to solve real world problem sets vital to national security.
  • Working knowledge with public keys and digital certificates.
  • Experience with DevOps environments.
  • Expertise in various COTS, GOTS, and open source tools which support development of data integration and visualization applications.
  • Experience with cloud message APIs and usage of push notifications.
  • Specialization in Object Oriented Programming languages, scripting, and databases.

Role Requirements :

  • Active TS/SCI
  • Full Time
  • High colocation, 70-80% onsite

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.clearancejobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:42 min

Comparing in-memory and Redis storage for cache scalability

Simone Sanfratello · World Congress 2022

Videos

See all

Related articles

See all