Full Stack ETL Developer

Fantom Corporation
Chantilly, VA, United States
14 days ago
Apply on www.clearancejobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours

Tech stack

Java (Programming Language) Agile Methodology Amazon Web Services Amazon S3 Apache HTTP Server Big Data C++ (Programming Language) Cloud Computing Cloud Engineering Continuous Integration Information Engineering Data Governance
+43 more
Extract Transform Load (ETL) Data Security Data Systems Distributed Computing Environment Amazon DynamoDB Protocol Buffers Graph Database JSON Python (Programming Language) Lightweight Directory Access Protocols (LDAP) PostgreSQL MariaDB MongoDB Monte Carlo Methods Neo4j NoSQL Cloud Services Next.js Scala (Programming Language) Unstructured Data WebGL Parquet Enterprise Software Applications Spring Cloud ReactJS Apache Spark Indexer Backend Pyspark Kubernetes Infrastructure Automation Frameworks Cassandra Avro AWS Fargate Apache Nifi Data Management Machine Learning Operations Restful APIs Stream Processing Data Pipelines Devsecops Docker Microservices

Job description

Fantom Corporation is a mission-focused organization supporting critical programs across the defense and intelligence community. We partner with our customers to deliver high-impact technical solutions while fostering a culture built on trust, expertise, and long-term career growth.

We are seeking a highly experienced Senior Data & Software Engineer to design, develop, and operate large-scale data platforms and mission-critical applications. This role combines data engineering, full-stack software development, cloud architecture, graph databases, MLOps, and DevSecOps.

The ideal candidate will have experience building highly scalable data systems, ETL/ELT pipelines, graph and NoSQL solutions, cloud-native applications, and modern user interfaces while operating within secure, regulated environments.

Responsibilities

  • Design and maintain enterprise-grade batch and real-time ETL/ELT pipelines
  • Develop front-end applications using React, Next.js, WebGL, or similar technologies
  • Develop backend services and microservices using Python, Java, Scala, C/C++, and REST APIs
  • Design and operate large-scale big data, NoSQL, relational, and graph database solutions
  • Build and optimize graph databases and traversal capabilities using technologies such as Gremlin, Cassandra, Neo4j, JanusGraph, and TinkerPop
  • Develop high-performance data processing pipelines using PySpark, Lambda, Step Functions, NiFi, and related technologies
  • Design data models, partitioning/sharding strategies, indexing, stream processing, and record aggregation workflows
  • Develop and maintain cloud-native solutions across AWS and other cloud platforms
  • Build and operate Kubernetes and Docker-based infrastructure
  • Develop CI/CD, Infrastructure as Code, DevSecOps, and MLOps pipelines
  • Implement data security, encryption, auditing, LDAP-based access controls, and data governance
  • Support federal security, compliance, and accreditation requirements
  • Collaborate across engineering and stakeholder teams to develop technical strategies that meet mission requirements

Requirements

  • Experience designing and operating large-scale data systems supporting billions to trillions of records/events
  • Strong experience with ETL/ELT, batch and real-time data pipelines, and distributed data processing
  • Full-stack development experience with React/Next.js and backend technologies such as Python, Java, or Scala
  • Experience with REST APIs and microservices architectures
  • Experience with Docker, Kubernetes, CI/CD, and Infrastructure as Code
  • Experience with AWS or other major cloud platforms
  • Strong experience with relational, NoSQL, and graph databases
  • Experience with technologies such as DynamoDB, Cassandra, PostgreSQL, MongoDB, Neo4j, ELK, MariaDB, MinIO, and S3
  • Experience with Spark/PySpark, Lambda, Step Functions, and stream-processing workflows
  • Knowledge of graph technologies such as Apache Gremlin, TinkerPop, and JanusGraph
  • Experience with probabilistic/statistical modeling, including risk scoring, Bayesian inference, and Monte Carlo simulation
  • Experience developing MLOps pipelines for large-scale applications
  • Experience with data security, governance, encryption, auditing, and LDAP
  • Experience implementing DevSecOps and Agile development practices in production environments
  • Experience with federal security, regulatory, compliance, and accreditation requirements
  • Experience working with structured, semi-structured, and unstructured data formats including JSON, CSV, AVRO, Parquet, and Protocol Buffers
  • CJ

Desired Qualifications

  • Experience designing and operating cloud-native data platforms within Intelligence Community environments
  • Experience with AWS ECS, Fargate, EMR, and multi-account AWS environments
  • Experience managing AWS Organizations, Organizational Units (OU), and Service Control Policies (SCP)
  • Experience with enterprise data catalogs, Policy Decision Points (PDPs), and data management services
  • Experience developing MLOps and large-scale data processing pipelines within secure government environments
  • Understanding of IT Service Management (ITSM) and SLA metrics
  • Experience presenting complex technical solutions and requirements to diverse technical and non-technical audiences
  • CJ

Requirements

  • Must be fully cleared with a recent polygraph
  • Must be willing and able to work fully onsite at the location listed in this posting

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.clearancejobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

2:24 min

Comparing Neo4j and GraphQL conceptual models

William Lyon · LIVE

3:02 min

Audience Q&A on data formats and engine tradeoffs

Matthias Niehoff Matthias Niehoff · World Congress 2026 Europe

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

3:30 min

Introduction to Neo4j and remote developer relations work

1:52 min

Customizing block storage tiers and formats

Ricardo Sueiras Sueiras · LIVE

Videos

See all

Related articles

See all