> Markdown version of [/jobs/ext/2010666-principal-data-scientist](https://www.wearedevelopers.com/jobs/ext/2010666-principal-data-scientist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Data Scientist - **Company:** Leonardo DRS - **Location:** Germantown, MD, United States - **Experience:** Expert - **Salary:** $134,332.0 - $181,349.0 - **Contract:** Permanent contract - **Skills:** Airflow, Data Analysis, Microsoft Azure, Data Transformation, Python (Programming Language), MATLAB, Machine Learning, SQL Databases, Delivery Pipeline, Apache Spark, Data Lineage, Data Pipelines, Databricks - **Published:** August 10, 2026 - **Apply:** https://careers.leonardodrs.com/talentcommunity/apply/1403829800/?locale=en_US ## About the Role * Master's degree in physical science or engineering discipline; PhD preferred * A minimum of 12 years of experience * Mastery of technologies and application domain * Expected to have knowledge and significant experience in multiple engineering disciplines * Must be a team player with good communication skills * Deep proficiency in Python for data analysis, preprocessing, and modeling * Strong foundation in applied statistics, analytics, and machine learning * Proven ability to set technical standards and mentor others, * Experience with anomaly detection, time-series analysis, and predictive monitoring. * Experience with data-pipeline and orchestration tooling (e.g., Airflow, Spark, SQL). * Exposure to geospatial and/or radar data concepts * Familiarity with Databricks and/or Azure * Experience building data-governance, lineage, or monitoring frameworks U.S. Citizenship required. This position requires an active DOD security clearance or the ability to obtain such clearance within a reasonable time after commencement of employment. ## Description * Own and build data-quality tooling, checks, and validation frameworks for time-series and geospatial datasets, ensuring datasets are ready for the ML training pipeline. * Lead classifier evaluation and validation. Define and apply rigorous metrics (precision/recall trade-offs, calibration, cross-validation, confusion-matrix and slice-based error analysis) and turn findings into concrete improvements. * Establish and enforce standards for data lineage, feature preparation (temporal alignment, resampling, lag features, spatial associations), and reproducibility. * Identify and resolve common ML data issues such as misaligned timestamps, invalid labels, class imbalance, and edge cases. * Learn and help refine our established model-training process; partner with the team to improve the end-to-end training pipeline. * Read, assess, and modernize legacy MATLAB tools; re-implementing in Python or other languages where it improves the process and results. * Mentor data engineers and scientists; document assumptions, methods, and limitations; and champion technical rigor across the team. * Provide briefings to customers, Senior Leadership, and internal stakeholders on capabilities and performance. * Able to work independently or lead small teams and drive projects to completion. * Work with customers to understand mission needs, review specifications and requirements, and develop solutions to best support them. * Provide budget, cost, and schedule input for design assignments. ## Related Videos - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Leveraging Large Language Models for Legacy Code Translation: Challenges and Solutions](https://www.wearedevelopers.com/videos/1157-leveraging-large-language-models-for-legacy-code-translation-challenges-and-solutions) - [Cutting LLM Costs Without Cutting Quality: How to Beat Proprietary LLMs with Fine-Tuned Open Source](https://www.wearedevelopers.com/videos/100151-cutting-llm-costs-without-cutting-quality-how-to-beat-proprietary-llms-with-fine-tuned-open-source) - [OLTP in the Lakehouse: Redefining Data for AI Workloads](https://www.wearedevelopers.com/videos/2038-oltp-in-the-lakehouse-redefining-data-for-ai-workloads) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Data Analyst Salary Germany](https://www.wearedevelopers.com/magazine/277-data-analyst-salary-germany) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)