Senior Data Scientist
Role details
Job location
Tech stack
Job description
- Exercises creativity in applying non-traditional approaches to large-scale analysis of unstructured data in support of high-value use cases visualized through multi-dimensional interfaces.
Requirements
-
Handle processing and index requests against high-volume collections of data and high-velocity data streams. Has the ability to make discoveries in the world of big data.
-
Requires strong technical and computational skills - engineering, physics, mathematics, coupled with the ability to code design, develop, and deploy sophisticated applications using advanced unstructured and semi-structured data analysis techniques and utilizing high-performance computing environments.
-
Has the ability to utilize advanced tools and computational skills to interpret, connect, predict, and make discoveries in complex data and deliver recommendations for business and analytic decisions.
-
Must be a U.S. Citizen
-
TS/SCI Clearance with polygraph
-
Minimum five (5) years of experience
-
Bachelor's degree in Computer Science or related field
-
Must be able to report daily to Tech Port San Antonio
-
Experience with software development, either an open-source enterprise software developmentstack (Java/Linux/Ruby/Python) or a Windows development stack (.NET, C#, C++).
-
Experience with hardware networking, server, OS (primarily Linux), GPU/DPU, and containerization (Docker primarily)
-
Experience NVIDIA Technology & frameworks
-
Experience with data transport and transformation APIs and technologies such as JSON, XML, XSLT, JDBC, SOAP and REST.
-
Experience with Cloud-based data analysis tools including Hadoop and Mahout, Acumulo, Hive, Impala, Pig, and similar.
-
Experience with visual analytic tools like Microsoft Pivot, Palantir, or Visual Analytics.
-
Experience with open-source textual processing such as Lucene, Sphinx, Nutch or Solr.
-
Experience with entity extraction and conceptual search technologies such as LSI, LDA, etc.
-
Experience with machine learning, algorithm analysis, and data clustering.
-
Ability to build Python scripts and packages that will be used by the data analysts. The Python scripts will use supervised and unsupervised ML (both text and image-based machine learning), natural language processing (named entity extraction, summarization, etc.), and network analysis (social network analysis, centrality analysis, dynamic network analysis).
-
Expertise in maintaining and deploying a notebook-based data science environment (JupyterHub).
-
Experience in advanced Python data science packages (Pandas, NetworkX, Scikit-Learn, PyTorch or TensorFlow/Keras, Matplotlib or Plotly, etc.)
Benefits & conditions
Compensation ranges encompass a total compensation package and are a general guideline only and not intended as a guaranteed and/or implied final compensation or salary for this job opening. Determination of official compensation or salary relies on several different factors including, but not limited to: level of position, complexity of job responsibilities, geographic location, candidate's scope of relevant work experience, educational background, certifications, contract-specific affordability, organizational requirements and alignment with local market data.
Our compensation includes other indirect financial components designed to support employees' total well-being, which should be considered when evaluating our competitive benefits package. These monetary benefits include medical insurance, life insurance, disability, paid time off, maternity/paternity leave, 401(k) company match, training/education reimbursements and other work/life programs.