Data Scientist
VTG LLC
Beavercreek, United States of America
4 days ago
Role details
Contract type
Permanent contract Employment type
Full-time (> 32 hours) Working hours
Regular working hours Languages
English Experience level
IntermediateJob location
Beavercreek, United States of America
Tech stack
Java
Artificial Intelligence
Amazon Web Services (AWS)
Amazon Web Services (AWS)
Apache HTTP Server
Applications Architecture
Application Integration Architecture
Azure
C++
Cloud Engineering
Information Systems
Computer Programming
Data Manipulation Languages
Data Visualization
Linux
DevOps
Programming Tools
HBase
Issue Tracking Systems
Python
Matlab
Machine Learning
MongoDB
NumPy
TensorFlow
Software Engineering
SQL Databases
Test Data
Unstructured Data
Jupyter Notebook
Google Cloud Platform
Data Storage Technologies
Spark
Keras
Gitlab
GIT
Pandas
Scikit Learn
Information Technology
Data Analytics
Dask
Kafka
Data Management
Data Pipelines
Jenkins
Artifactory
Job description
The Data Scientist will work with a team of DevOps engineers, software developers, data engineers, and system operators to identify data needs and prototype a range of novel solutions. This data scientist would be involved at all levels of the data life cycle from onboard management of data to its use in application development and back to application integration and gathering test data.
- Leverage third-party tools to architect and prototype a modern data management and application development pipeline in a local and/or a cloud environment
- Perform data analytics of simulated and real-world data
- Integrate structured and unstructured data from disparate data sources
- Develop applications and models supporting various users
- Provide technical input to program managers and government representatives
Requirements
- Bachelor's Degree, majoring in majoring in Computer Science, Data Science, Information Systems, or a related field
- 4+ years of experience as a Data Scientist including experience in statistical modeling and machine learning based on the analysis of large sets of data
- Experience with data storage and management tools (S3, SQL, MongoDB, Hbase, Apache Atlas, Kafka, etc.)
- Programming experience in Python, R, or similar data manipulation languages and associated libraries (e.g. pandas, numpy, polars, dask)
- Experience with data science and analytics toolsets (e.g. JupyterHub / Jupyter Notebooks, Apache Spark, MATLAB)
- Knowledge of data modeling principles
- Experience in knowledge extraction and insights from data in various forms, both structured and unstructured
- Cloud development experience, preferably in AWS
- Excellent verbal and written communication skills
- US Citizen with current TOP SECRET/SCI Eligible Clearance or ability to obtain a TOP SECRET/SCI clearance
- Successful completion of background check
Desired qualifications:
- Master's Degree or higher in Computer Science, Data Science, or Information Systems
- Experience establishing data pipelines in cloud platforms, such as AWS, Azure, or Google Cloud
- Data visualization experience and associated tools/libraries (e.g. pyplot, seaborn)
- Experience using Git for version control and issue tracking
- Experience with artifact repositories (e.g. Artifactory)
- Experience with CI/CD pipelines (e.g. Jenkins, Gitlab pipelines)
- Experience with AI/ML development tools and libraries (e.g. Sagemaker, ML Studio, Tensorflow, Keras, scikit-learn)
- Experience leading teams and projects
- Programming experience in C++ and Java
- Experience with Linux systems
About the company
VTG is seeking a Data Scientist to support our Team in Beavercreek, OH. The Data Scientist will design, prototype, and implement a data management and application development pipeline in support of national defense data science and data architecture prototyping tasks. This role will also include gathering and organizing data, conducting data analytics, and developing data analytic and AI/ML based applications. This is an onsite role due to its classification level.