Forward Deployed Data Engineer

BOAB Ventures
Groton, CT, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Artificial Intelligence Airflow Amazon S3 Apache HTTP Server Batch Processing Databases Continuous Integration Information Engineering Data Warehousing Distributed Systems Github Protocol Buffers
+16 more
Python (Programming Language) PostgreSQL Message Queuing Telemetry Transport (MQTT) Online Analytical Processing Online Transaction Processing Parsing Data Streaming Systems Integration Parquet Data Processing SC Clearance Apache Kafka Data Management Software Coding Stream Processing Data Pipelines

Job description

We’re seeking a skilled Forward Deployed Data Engineer to help build next-generation data management and artificial intelligence platforms supporting maritime domain awareness missions.

As a forward-deployed engineer, you’ll work directly with customers-writing code, integrating systems, and solving complex technical challenges in operational environments. You’ll serve as both an engineer and technical advisor, taking ownership of mission outcomes. This position requires a current or active U.S. Secret clearance.

This role offers the opportunity to work on real-world projects involving sensor systems, edge data collection, and large-scale maritime data processing environments that directly support operational mission success.

Responsibilities

  • Implement real-time data pipelines using MQTT and Redpanda for stream processing.
  • Implement offline data pipelines using Dagster for batch processing.
  • Parse and process binary message formats from various data sources.
  • Build data warehouses using PostgreSQL, Apache Iceberg, Parquet, and S3.
  • Design data models that support high-performance analytical queries.
  • Validate and normalize diverse data sources.
  • Improve local development environments and CI/CD workflows using modern tooling and GitHub Actions.

Requirements

  • Expertise in time-series data processing and analysis (windowing, resampling, interpolation, etc.)
  • Proficiency in Python and Rust for data engineering workflows
  • Experience with binary message parsing
  • Experience with row-based and columnar data formats
  • Experience with OLTP and OLAP databases
  • Knowledge of distributed systems, streaming architectures, and batch processing patterns
  • Hands-on experience with workflow orchestration platforms such as Dagster or Airflow
  • Hands-on experience with streaming platforms such as Redpanda or Kafka
  • Experience working with binary message formats such as Protobuf

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.clearancejobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

2:50 min

How Parquet metadata enables efficient data reading

Matthias Niehoff Matthias Niehoff · WWC Europe 2026

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:03 min

Introduction to open table formats built on Parquet

Matthias Niehoff Matthias Niehoff · WWC Europe 2026

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all