> Markdown version of [/jobs/ext/1951473-data-scientist-nlp](https://www.wearedevelopers.com/jobs/ext/1951473-data-scientist-nlp). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist - NLP - **Company:** ANALYTICA - **Location:** United States (Remote available) - **Experience:** Experienced - **Contract:** Temporary contract - **Skills:** Artificial Intelligence, Amazon Web Services, Artificial Neural Networks, Microsoft Azure, Data Files, Python (Programming Language), Machine Learning, Machine Translation, Natural Language Processing, Named Entity Recognition, NLTK (NLP Analysis), Open Source Technology, Tensorflow, SAS (Software), Sentiment Analysis, Stemming, Feature Engineering, Pytorch, Large Language Models, Prompt Engineering, Apache Spark, Deep Learning, Topic Modeling, Keras, Information Technology, Machine Learning Operations, Gensim, Spacy, Software Version Control, Unsupervised Learning, Databricks - **Published:** August 6, 2026 - **Apply:** http://analyticallc.applytojob.com/apply/jobs/details/XnxisVWUot ## About the Role * Feature Engineering and Attribute Evaluation - Candidate must demonstrate experience with NLP feature engineering methods such as TF-IDF, word2vec, GloVe, and FastText identifying the key determinants for modeling that exist in the business process and within existing data sets as well as selecting evaluation protocols (model techniques). * Modeling - Candidates will have practiced skills and experience selecting classification modeling techniques to fit the business problem. Examples will include techniques such as machine learning (ML) supervised and unsupervised learning, regression, neural networks and deep learning, natural language processing, etc. * Validation - Strong candidates will describe their experience with investigating, reporting, and justifying model results. * Visualization- Experience in presenting the results of their modeling activities, depicting the insights realized, and explaining the relevance of their results to the organization's business challenges., * Master's degree required, and PhD preferred in Statistics, Mathematics, Computer Science, or similar * High degree of experience utilizing SAS, R, or Python to support NLP use cases such as Document Summarization, Named Entity Recognition, Sentiment Analysis, and/or Topic Modeling * At least four years of experience developing scalable, production-ready NLP solutions using sci-kit learn, Keras, TensorFlow, PyTorch, Spark NLP. * Experience using git/github to version control source code * Experience leveraging transformer architecture to develop NLP models * Experience with open source NLP packages such as Gensim, SpaCy, or NLTK. * Experience with BERT, GPT-J, RoBERTa, T5 or other transformers * Experience with GenAI and Prompt Engineering is a plus * Experience in Databricks and MLFlow is a plus * Experience with machine translation and transcription of foreign language documents using Microsoft Azure translation services is a plus * Experience working in an AWS cloud environment and with related AWS services such as Bedrock and Textract * Experience coordinating and maintaining user stories * Must be a US citizen * Must be able to obtain and maintain a Public trust security clearance ## Description * Pre-processing - Demonstrate the skills and experience to collect, clean, and prepare data sets for input into a computational model using Python. Strong candidates will explain various methods you have applied using common pre-processing functions such as stop word removal, stemming, lemmatization, and tokenization., To enhance efficiency, fairness, and accuracy, Analytica may use AI-assisted tools to support certain aspects of our hiring process. * Application Review: AI tools may help identify skills and experiences relevant to the role. * Interview Support: AI-powered notetaking tools may be used during interviews to document discussions and summarize key points. These tools are used to assist our team. All hiring decisions are made by Analytica recruiters and hiring managers. ## Related Videos - [A beginner’s guide to modern natural language processing](https://www.wearedevelopers.com/videos/858-a-beginner-s-guide-to-modern-natural-language-processing) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [Overview of Machine Learning in Python](https://www.wearedevelopers.com/videos/840-overview-of-machine-learning-in-python) - [Outclassing Frontier LLMs at Extracting Information](https://www.wearedevelopers.com/videos/100303-outclassing-frontier-llms-at-extracting-information) - [AI That Fits Your Business, Not the Other Way Around](https://www.wearedevelopers.com/videos/100148-ai-that-fits-your-business-not-the-other-way-around) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)