Forward Deployed Data Engineer

BOAB Ventures
Groton, United States of America
21 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English

Job location

Groton, United States of America

Tech stack

Artificial Intelligence
Airflow
Amazon Web Services (AWS)
Apache HTTP Server
Batch Processing
Databases
Continuous Integration
Information Engineering
Data Warehousing
Distributed Systems
Github
Protocol Buffers
Python
PostgreSQL
Message Queuing Telemetry Transport (MQTT)
Online Analytical Processing
Online Transaction Processing
Parsing
Data Streaming
Systems Integration
Parquet
Data Processing
SC Clearance
Kafka
Data Management
Software Coding
Stream Processing
Data Pipelines

Job description

We're seeking a skilled Forward Deployed Data Engineer to help build next-generation data management and artificial intelligence platforms supporting maritime domain awareness missions.

As a forward-deployed engineer, you'll work directly with customers-writing code, integrating systems, and solving complex technical challenges in operational environments. You'll serve as both an engineer and technical advisor, taking ownership of mission outcomes. This position requires a current or active U.S. Secret clearance.

This role offers the opportunity to work on real-world projects involving sensor systems, edge data collection, and large-scale maritime data processing environments that directly support operational mission success.

Responsibilities

  • Implement real-time data pipelines using MQTT and Redpanda for stream processing.
  • Implement offline data pipelines using Dagster for batch processing.
  • Parse and process binary message formats from various data sources.
  • Build data warehouses using PostgreSQL, Apache Iceberg, Parquet, and S3.
  • Design data models that support high-performance analytical queries.
  • Validate and normalize diverse data sources.
  • Improve local development environments and CI/CD workflows using modern tooling and GitHub Actions.

Requirements

  • Expertise in time-series data processing and analysis (windowing, resampling, interpolation, etc.)
  • Proficiency in Python and Rust for data engineering workflows
  • Experience with binary message parsing
  • Experience with row-based and columnar data formats
  • Experience with OLTP and OLAP databases
  • Knowledge of distributed systems, streaming architectures, and batch processing patterns
  • Hands-on experience with workflow orchestration platforms such as Dagster or Airflow
  • Hands-on experience with streaming platforms such as Redpanda or Kafka
  • Experience working with binary message formats such as Protobuf

Apply for this position