> Markdown version of [/jobs/ext/1957433-data-scientist-nlp](https://www.wearedevelopers.com/jobs/ext/1957433-data-scientist-nlp). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist - NLP - **Company:** ANALYTICA - **Location:** Washington, DC, United States (Remote available) - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Data Cleansing, Python (Programming Language), Machine Learning, Natural Language Processing, NLTK (NLP Analysis), Open Source Technology, Tensorflow, SAS (Software), Feature Engineering, Pytorch, Large Language Models, Deep Learning, Model Validation, Topic Modeling, Keras, Scikit Learn, Information Technology, Performance Monitor, Machine Learning Operations, Gensim, Spacy, Software Version Control, Databricks - **Published:** August 6, 2026 - **Apply:** https://us.experteer.com/career/view-jobs/data-scientist-nlp-washington-dc-usa-58817339 ## About the Role The or similar * 4+ years of experience developing scalable, production-ready NLP solutions * Proficiency in SAS, R, or Python for NLP tasks (summarization, NER, sentiment, topic modeling) * Experience with scikit-learn, Keras, TensorFlow, PyTorch; transformer architectures * Experience with Git/GitHub for version control * Familiarity with open-source NLP packages (Gensim, SpaCy, NLTK) and transformers (BERT, GPT-J, RoBERTa, T5) * Experience with Databricks and MLFlow is a plus * Experience with AWS services (Bedrock, Textract) is a plus * Experience with Azure translation services or similar is a plus * US citizenship and ability to obtain Public Trust security clearance Key requirements * competitive compensation * employer-paid health care * training and development funds * 401k match * bonus opportunities ## Description Experteer Overview In this Data Scientist role, you apply statistical programming, modeling, and NLP to tackle public sector challenges for federal client engagements. You will work on data preparation, feature engineering, modeling, validation, and visualization to translate complex data into actionable insights. You'll collaborate with cross-functional teams to address mission-critical problems in health, civilian, and national security domains. The opportunity offers remote work, strong growth potential, and the chance to contribute to impactful government-facing analytics. Compensation / Benefits * Data pre-processing and cleaning with Python for model-ready inputs * NLP feature engineering using TF-IDF, word2vec, GloVe, FastText * Developing classification, ML, deep learning, and NLP models * Model validation, justification, and performance reporting * Visualizing results and communicating insights to stakeholders Tasks * Master's degree in Statistics, Mathematics, Computer Science, or similar * 4+ years of experience developing scalable, production-ready NLP solutions * Proficiency in SAS, R, or Python for NLP tasks (summarization, NER, sentiment, topic modeling) * Experience with scikit-learn, Keras, TensorFlow, PyTorch; transformer architectures * Experience with Git/GitHub for version control * Familiarity with open-source NLP packages (Gensim, SpaCy, NLTK) and transformers (BERT, GPT-J, RoBERTa, T5) * Experience with Databricks and MLFlow is a plus * Experience with AWS services (Bedrock, Textract) is a plus * Experience with Azure translation services or similar is a plus * US citizenship and ability to obtain Public Trust security clearance Key requirements * competitive compensation * employer-paid health care * training and development funds * 401k match * bonus opportunities ## Related Videos - [A beginner’s guide to modern natural language processing](https://www.wearedevelopers.com/videos/858-a-beginner-s-guide-to-modern-natural-language-processing) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Overview of Machine Learning in Python](https://www.wearedevelopers.com/videos/840-overview-of-machine-learning-in-python) - [Serverless deployment of (large) NLP models ](https://www.wearedevelopers.com/videos/158-serverless-deployment-of-large-nlp-models) - [Introduction to Azure Machine Learning](https://www.wearedevelopers.com/videos/368-introduction-to-azure-machine-learning) ## Related Articles - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk)