> Markdown version of [/jobs/ext/3054235-data-scientist](https://www.wearedevelopers.com/jobs/ext/3054235-data-scientist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist - **Company:** Geo Owl - **Location:** United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), JavaScript (Programming Language), Geographic Information Systems, Application Programming Interfaces (APIs), Artificial Intelligence, Computer Vision, Cloud Engineering, Computer Programming, Databases, Data Architecture, Information Engineering, Data Governance, Data Infrastructure, Extract Transform Load (ETL), Data Transformation, Data Migration, Data Structures, Data Systems, Identity and Access Management, Python (Programming Language), NoSQL, Operational Data Store, Standard Sql, DataOps, Scaled Agile Framework, SQL Databases, Enterprise Data Management, Enterprise Software Applications, Cloud Platform System, Backend, Machine Learning Operations, Api Design, Restful APIs, Data Pipelines, Programming Languages - **Published:** September 24, 2026 - **Apply:** https://www.thejobnetwork.com/job/d3cd78ed-91bf-4818-b8c7-360f9740773d/senior-data-scientist ## About the Role * *Active TS/SCI clearance with eligibility for SI/TK/G/HCS/NATO access; CI polygraph required or willingness to obtain * *Minimum 10 experience points required (see experience point calculation below) * *4+ years of experience working with SQL or NoSQL databases and cloud architectures * *Experience leveraging APIs to query systems and migrate data * *Experience with programming languages including Python, JavaScript, or Java * *Experience developing and sustaining operational data pipelines supporting multiple data formats and interfaces * *Excellent written and verbal communication skills; ability to articulate complex technical concepts to both lay and expert audiences Preferred Qualifications * *Experience with government enterprise technologies, authorization processes, and secure government system access management * *Experience implementing processes and tools within an organization using the Scaled Agile Framework for Enterprise (SAFe) * *Experience developing dynamic visualizations to monitor pipeline availability, currency, and volumetrics Experience Point Requirement This is a Senior-level (Level 4) position requiring a minimum of 10 experience points. Points are calculated as follows: * Education: Associate's = 2 pts · Bachelor's = 3 pts · Master's = +2 pts · PhD = +3 pts * Professional / Military Experience: 1 pt per year of relevant experience * Certifications: 0.5 pts each * Specialized Training: 0.25 pts per relevant course * Professional Impact (publications, presentations, patents): up to 3 pts total ## Description The imagery that supports the Government's AI model evaluation program doesn't move itself. It has to be ingested, standardized, transformed, and routed across security domains with precision - and that requires a data engineer who can build the backend that makes it all work. As a Senior Data Scientist on this program, you'll own the data architecture and pipeline infrastructure that enables the accreditation pipeline to run: building ETL workflows, managing database schemas, and solving the cross-domain data problems that stand between an AI model and its accreditation clearance., As a Senior Data Scientist on the program, you support the maintenance and implementation of the data architecture that enables the program to store, process, and move imagery data across security domains and systems. Working alongside imagery scientists and the government customer team, you'll build and sustain the pipelines and database infrastructure needed to ingest emergent sensor data, apply ETL transformations to standardize it for labeling and model testing, and move it between platforms via API. You'll bring deep database and cloud architecture experience, strong programming skills, and the ability to work within classified government IT environments. This is a SCIF-based role performed at a contractor facility., The Data Scientist shall support the maintenance and implementation of the data architecture for all Program data. This position requires experience with databases, cloud architectures, and compliance with data standards. This position is responsible for maintaining data backend workflows that allow for the movement of data across security domains and across systems. A Day in the Life * *Build and maintain database tables, schemas, and health alerts for emergent sensor data; set up ETL pipelines to ingest, transform, and load new imagery data in alignment with government schemas and governance requirements * *Develop API-based workflows to enable data movement between platforms as directed by the government; assess metadata, format, and schema differences when integrating new sensors into the existing data operations pipeline * *Determine pre-processing and standardization approaches for new sensor data - file format conversions, geospatial tiling (chipping), orthorectification, and other data transformations needed for labeling and model testing * *Monitor data health and pipeline availability; address issues that could block imagery delivery to the accreditation testing pipeline * *Ensure all sensor data adheres to government-directed schemas, formats, and governance requirements; maintain documentation and standards compliance throughout the pipeline * *Available during core business hours (10am-2pm EST) to coordinate with imagery scientists and the government customer on data architecture and pipeline questions Why This Role Matters Mission Impact program cannot run AI model evaluation and testing testing without properly formatted, schema-compliant imagery data that has been moved securely across the domains where it needs to live. Your pipeline work is the infrastructure that connects imagery identification to model testing - and when it works, AI models get accredited and cleared for operational use. Geo Owl Impact Your engineering reliability is a Key Position deliverable on this program. Your work sustains Geo Owl's technical standing on program and demonstrates the data capability that makes us a credible partner on this contract. Your Growth You'll develop deep expertise in classified government cloud and data pipeline environments - including cross-domain data solutions, enterprise data platforms and MLOps technologies, and AI/ML-ready data infrastructure - in a program context where data engineering directly enables national security outcomes. Core Responsibilities * *Build and maintain database infrastructure (tables, schemas, health alerts, ETL pipelines) to store and process emergent sensor data in support of program AI model evaluation and testing * *Develop API-based workflows to enable data movement between platforms across security domains as directed by the government * *Ensure all sensor data adheres to government-directed data schemas, formats, and governance requirements throughout the pipeline * *Assess metadata, format, and data structure differences for new sensors; adapt ETL processes, schemas, and APIs accordingly to enable clean ingestion into the existing data operations pipeline * *Determine pre-processing and standardization requirements for new sensor data - file format conversion, tiling/chipping to specified geospatial bounds, orthorectification, and other data transformations required for labeling and model testing * *Maintain pipeline reliability and data availability across program's operational environment; respond to data health issues that could block the accreditation testing workflow, Tools, Technologies & Tradecraft Python JavaScript / Java SQL NoSQL Cloud Architecture REST APIs ETL Pipelines Geospatial Data Formats Orthorectification / Chipping Government Enterprise Technologies [Preferred] SAFe [Preferred] What Makes Someone Successful in This Role * *You're fluent across the data stack - schema design, API development, ETL orchestration, and cross-domain data movement feel like one integrated problem to you, not separate tracks * *You navigate classified environments with precision - you understand that government system access credentials and ATO requirements, and security domain constraints aren't overhead, they're the architecture * *You can assess a new sensor's metadata and schema against an existing pipeline and tell the team exactly what needs to change before the first data arrives * *You take governance seriously - you understand that non-compliant data in the accreditation pipeline produces unreliable test results, and you prevent it at the ETL layer * *You communicate your engineering decisions clearly - you can explain a schema change or a pipeline failure to imagery scientists and program managers who don't speak SQL Is This Role For You? Great Fit You are a data engineer who thrives in structured, high-compliance classified environments and wants your pipeline work to matter at the national security level. You have experience with government cloud infrastructure or cross-domain data solutions and are energized by the challenge of keeping complex, multi-sensor data operations running in a SCIF-based environment. May Not Be For You This may not be the right fit if you prefer rapid-iteration commercial cloud environments without classified IT constraints, or if working from a certified contractor SCIF rather than remotely from home is not compatible with your situation. Career Growth & Professional Value This role builds deep expertise in geospatial data engineering within a classified government environment - including cross-domain data solutions, enterprise data platforms, and AI/ML-ready pipeline architecture. Experience maintaining data infrastructure for a government AI model evaluation program that spans multiple computer vision programs (multiple AI/ML programs) is a highly differentiated credential for data engineers pursuing senior or architect-level roles in the defense and intelligence community. ## Related Videos - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Cyber Sleuth: Finding Hidden Connections in Cyber Data](https://www.wearedevelopers.com/videos/893-cyber-sleuth-finding-hidden-connections-in-cyber-data) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Data Science, ML & AI in the Oil and Gas Industry at NDT Global - Dr. Katja Träumner](https://www.wearedevelopers.com/videos/1308-data-science-ml-ai-in-the-oil-and-gas-industry-at-ndt-global-dr-katja-traumner) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) - [NoSQL Data Modeling for Front-end Developers](https://www.wearedevelopers.com/videos/297-nosql-data-modeling-for-front-end-developers) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story)