Data Engineer - Spear AI

Spear AI
Washington, DC, United States
4 days ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Airflow Amazon Web Services Cloud Computing Security Information Engineering Data Governance Data Security Decision Support Systems Python (Programming Language) Network Security SQL Databases Data Processing Data Ingestion
+4 more
Apache Spark SC Clearance Data Management Data Pipelines

Job description

Spear AI seeks a Data Engineer to build robust, secure data infrastructure supporting the an IC Task Force. The role involves creating resilient pipelines to power analytics and decision support systems. Primary Responsibilities: Design and manage secure data pipelines enabling ML and analytics workflows. Handle data ingestion, transformation, and validation with auditable, secure processes. Collaborate with mission engineers and data scientists to align infrastructure with operational objectives. Optimize data processing and storage for real-time and batch operations. Ensure compliance with DIA’s data governance and cross-domain security standards.

Requirements

Active Secret clearance. 3-7 years in data engineering, preferably within secure, high-side environments. Proficiency in Python, Spark, SQL, and data orchestration tools (e.g., Airflow). Experience with classified data management, secure networking, and infrastructure optimization. Preferred Experience: Familiarity with IC standards (UDS, IC ITE) and secure cloud environments (AWS GovCloud, C2S). Strong troubleshooting and optimization skills within complex operational settings.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

2:04 min

Comparing offline data analytics with online stream processing

Artem Volk Artem Volk +1 · World Congress 2024

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all