> Markdown version of [/jobs/ext/2598091-senior-data-scientist-machine-learning](https://www.wearedevelopers.com/jobs/ext/2598091-senior-data-scientist-machine-learning). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Scientist - Machine Learning - **Company:** General Dynamics Information Technology - **Location:** Fairfax, VA, United States (Remote available) - **Experience:** Expert - **Salary:** $123,250.0 - $166,750.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Network Analysis, Databases, Data Warehousing, Fraud Prevention and Detection, Python (Programming Language), Machine Learning, SQL Databases, Management of Software Versions, Supervised Learning, Feature Engineering, Snowflake, Information Technology, Machine Learning Operations - **Published:** August 16, 2026 - **Apply:** https://dejobs.org/x/x/7F33728CFF7E44C391B15BC95DFDF7A7/job/ ## About the Role Amazon Web Services (AWS),Healthcare Claims,Predictive Modeling,Python (Programming Language),Supervised Learning Experience: 5 + years of related experience, * Master's in a quantitative field (statistics, computer science, engineering, applied mathematics or related), or a Bachelor's with equivalent hands-on experience. * 5+ years building, validating and delivering supervised machine learning models on real-world data, including work in which labels were incomplete, delayed or biased. * Experience deploying models into production and maintaining them: scheduling execution, versioning and drift monitoring, with data engineers. * Python and SQL, including feature engineering within the data warehouse at very large scale rather than extracting to a local environment. * 2+ years working with healthcare claims data (Medicare, Medicaid or commercial) and coding systems (e.g., ICD-10, CPT, HCPCS, DRG). * Experience with validation design for imbalanced, temporally shifting problems: out-of-time evaluation, leakage detection, calibration and precision-focused metrics over ranked output. * Ability to explain model output to a non-technical investigator, defend methodology to technical audiences, and present analytic outcomes to clients and stakeholders., * Graph or network analytics, entity resolution and record linkage for identifying collusive relationships across payers. * Positive-unlabeled, semi-supervised or active learning against a capacity-constrained review queue. * Modeling in a regulated or adverse-action setting where explainability and fairness were requirements. * Anomaly detection, peer-group construction and case-mix methods (e.g., HCC); AWS and/or Snowflake, including Snowpark or model lifecycle tooling. * Healthcare FWA or program integrity datamining in multi-payer databases; payer coverage policy (LCDs, NCDs) and claim edits (e.g., NCCI). SECURITY CLEARANCE LEVEL: * Must be able to obtain/maintain Public Trust. ## Description As the Senior Data Scientist for Machine Learning supporting the Healthcare Fraud Prevention Partnership (HFPP), you will be the first dedicated machine learning practitioner at the Trusted Third Party (TTP), an established Fraud, Waste and Abuse (FWA) analytics program. You will develop predictive models against a multi-billion record claims warehouse assembled from dozens of public and private healthcare payers, and you will establish how machine learning models move from development into production on this program. The data, the subject matter experts and the payer partnerships are already in place; the modeling capability is yours to build. This is a senior individual contributor position without direct reports, and it is the only role on the team focused primarily on machine learning, meaning the Senior Data Scientist will be establishing practice rather than joining one. ***Work visa sponsorship will not be provided for this position. This is a remote role. Candidates must reside in the United States. MEANINGFUL WORK AND PERSONAL IMPACT: * Designing, training and validating supervised models that score providers and billing patterns for FWA risk, using investigative case-level data, payer feedback on referred leads, and public exclusion and enforcement data as labels, including the entity resolution to link enforcement records to providers in claims. * Designing validation for the actual conditions: labels lagging billing behavior by years, coverage limited to leads previously referred, extreme class imbalance, and schemes that shift faster than confirmation arrives. * Engineering features against billions of claim records within the warehouse rather than extracting data to local memory, using Python and SQL, alongside data engineers and Business Intelligence Developers. * Delivering output that supports action. Investigators need the specific claims, the pattern and the basis for the finding, so each model carries a human-readable rationale and claim-level evidence alongside the score, adjusted for case mix and specialty and ranked so that precision at the top of the review queue is the operative measure. * Deploying models into production and keeping them healthy, including scheduled execution, versioning and drift monitoring, and establishing the modeling and deployment practices the Data Science team adopts going forward. * Collaborating with FWA Subject Matter Experts to separate genuine anomalies from patterns explained by coverage policy or claim edits, and communicating methodology and limitations to HFPP Partners and stakeholders so that output is adopted and acted upon. ## Related Videos - [Fault Tolerance and Consistency at Scale: Harnessing the Power of Distributed SQL Databases](https://www.wearedevelopers.com/videos/1146-fault-tolerance-and-consistency-at-scale-harnessing-the-power-of-distributed-sql-databases) - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) - [Kubernetes and Microservices with Multi-Model Databases](https://www.wearedevelopers.com/videos/382-kubernetes-and-microservices-with-multi-model-databases) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Fault Tolerance and Consistency at Scale: Harnessing the Power of Distributed SQL Databases](https://www.wearedevelopers.com/videos/1520-fault-tolerance-and-consistency-at-scale-harnessing-the-power-of-distributed-sql-databases) - [Beyond SQL Generation: How to Teach Agents What Your Database Actually Means](https://www.wearedevelopers.com/videos/100127-beyond-sql-generation-how-to-teach-agents-what-your-database-actually-means) ## Related Articles - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)