Data Scientist

NAVON TECHNOLOGIES LLC
Chantilly, VA, United States
28 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Data Analysis Cloud Computing Data Files Relational Databases Github Python (Programming Language) Machine Learning Language Modeling NLTK (NLP Analysis) Tensorflow Search Technologies
+19 more
SQL Databases Tableau (Software) Web Applications Scripting Graphics Processing Unit (GPU) Pytorch Flask (Web Framework) Large Language Models Deep Learning Software Application Programming Topic Modeling Generative AI Keras Scikit Learn HuggingFace Gensim Spacy Document Classification Jenkins

Requirements

Demonstrate professional or academic experience performing NLP tasks, including selecting the best Python libraries for a given task, choosing appropriate pre-processing actions, performing analysis, and assessing model performance.

Demonstrate professional or academic experience using Python NLP packages such as Spacy, Gensim, or NLTK to analyze or process collections of documents.

Demonstrate professional or academic experience with deep learning frameworks such as PyTorch, Tensorflow, or Keras.

Demonstrate professional or academic experience with the HuggingFace Transformers library and hub.

Demonstrate experience creating machine learning models that conduct text classification and topic modeling in Python using standard machine learning (Scikit-learn) or deep learning models.

Demonstrate academic or professional experience using encoder-decoder and generative language models to perform NLP tasks.

Demonstrate academic or professional experience communicating methodological choices and model results.

Demonstrate professional or academic experience and proficiency with SQL to include using common table expressions, set operations, aggregated functions and nested subqueries.

Demonstrate professional or academic experience with version control systems such as Github and Jenkins.

Demonstrate experience leveraging GPUs for accelerated computing.

Develop practical approaches for measuring performance.

Assist in developing types of measure, the collection of data, analyzing the data, and presenting that data to senior leadership.

Conduct advanced statistical analysis on personnel, intelligence and performance metrics.

Assist in selection or development of appropriate methodology to conduct research.

Analyze information and provide research findings in a manner that is easily grasped by the customer and consumers.

Desired Skills:

Experience writing Python scripts that pull data from web-based APIs and relational databases.

Experience with cloud computing development and architecture.

Experience with front-end web development frameworks such as Flask.

Experience developing applications for semantic search.

Experience tuning LLMs on custom data sets and applying results to specific use cases.

Demonstrate professional or academic experience and proficiency with Tableau to produce visualizations and dashboards.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.clearancejobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:13 min

Training a Word2Vec model with Python and Gensim

Jodie Burchell · LIVE

1:18 min

Converting existing Keras models to TensorFlow format

Håkan Silfvernagel · LIVE

3:29 min

Binary and count vectorization techniques for text

Jodie Burchell · LIVE

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

2:20 min

Speeding up model training cycles with transfer learning techniques

Anirudh Koul · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all