> Markdown version of [/jobs/ext/2405270-senior-data-scientist](https://www.wearedevelopers.com/jobs/ext/2405270-senior-data-scientist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Scientist - **Company:** Inspyr Solutions - **Location:** Woodlawn, MD, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Big Data, Code Review, Encodings, Continuous Integration, Data Cleansing, Data Security, IBM DB2, Text Processing, Apache Hadoop, Information Extraction, Information Sciences, Python (Programming Language), PostgreSQL, Microsoft SQL Server, Natural Language Processing, Named Entity Recognition, Oracle (Applications), Regular Expressions, SQL Databases, Unstructured Data, Freeform SQL, Indexer, Scikit Learn, Information Technology, Spacy, Software Version Control, Data Pipelines, Jenkins - **Published:** August 18, 2026 - **Apply:** https://www.dice.com/job-detail/1fe9686d-c01d-4065-9aa8-b7b250b4b9cb ## About the Role * Bachelor's degree in Computer Science, Statistics, Applied Mathematics, Information Science, or related field. * 10+ years of IT experience; alternatively, Master's + 10 years, Bachelor's + 12 years, or 18+ years of relevant experience in lieu of a degree. * Strong hands-on experience with NLP, text processing, information extraction, Python, SQL, and Regex. * Experience with Named Entity Recognition, entity resolution, record linkage, and data matching techniques. * Experience with Python libraries/frameworks such as spaCy and Scikit-learn. * Familiarity with record linkage tools such as Splink, FastLink, Dedupe, or recordlinkage. * Excellent written and verbal communication skills. * Ability to obtain and maintain a Public Trust clearance. * Must be willing to work on-site 5 days per week in Woodlawn, MD. Preferred Qualifications * Experience supporting federal, state, or local government data initiatives. * Experience with PostgreSQL, DB2, Oracle, SQL Server, Hadoop, and flat files. * Experience with Jenkins, CI/CD, and data pipeline automation. * Experience designing and owning data pipeline architectures from discovery through implementation and monitoring. * Strong analytical, problem-solving, and cross-functional collaboration skills. ## Description We are seeking a Senior Data Scientist with strong experience in NLP, text processing, information extraction, entity resolution, Python, and SQL. This role will design and optimize data processing pipelines, work with large and complex datasets, and develop scalable analytics solutions in an enterprise environment., * Design, develop, and maintain data processing and entity resolution pipelines using Python and SQL. * Apply NLP and text-processing techniques to clean, match, standardize, and extract information from structured and unstructured data. * Develop solutions using Named Entity Recognition, TF-IDF/Cosine Similarity, Blocking/Indexing, String Distance Metrics, Phonetic Encoding, and Address Standardization. * Optimize complex SQL queries and database operations for performance and scalability. * Perform data cleansing, transformation, validation, testing, deployment, and monitoring. * Participate in code reviews and follow version control, data security, and reproducibility best practices. * Communicate complex technical concepts and analytical decisions effectively to technical and non-technical stakeholders. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [The Road to MLOps: How Verivox Transitioned to AWS](https://www.wearedevelopers.com/videos/1050-the-road-to-mlops-how-verivox-transitioned-to-aws) - [Optimizing Discovery: PostgreSQL's Role in Transforming GetYourGuide's Search](https://www.wearedevelopers.com/videos/1647-optimizing-discovery-postgresql-s-role-in-transforming-getyourguide-s-search) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) - [Our GitOps approach for deploying an Identity Provider and an API Gateway in a SaaS company](https://www.wearedevelopers.com/videos/776-our-gitops-approach-for-deploying-an-identity-provider-and-an-api-gateway-in-a-saas-company) - [Dynamic Entities in .NET: Building Low-Code Systems on Top of Entity Framework Core](https://www.wearedevelopers.com/videos/100218-dynamic-entities-in-net-building-low-code-systems-on-top-of-entity-framework-core) ## Related Articles - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)