Senior Data Scientist

Master Compliance
Syracuse, NY, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$83,934.0 - $101,082.0
Working hours
Regular working hours
Job source

Tech stack

C (Programming Language) Java (Programming Language) Artificial Intelligence Amazon Web Services Bash Shell Big Data Computer Programming Data Cleansing Data Integration Extract Transform Load (ETL) Data Mining Data Visualization
+26 more
Database Design R (Programming Language) Apache Hadoop Hadoop Distributed File System Python (Programming Language) Unix Shell Machine Learning Natural Language Processing NoSQL Tensorflow SAS (Software) SQL Databases Talend Unstructured Data Data Storage Technologies Feature Engineering Apache Spark Model Validation Pandas Spark Mllib Scikit Learn Vba Programming Language Machine Learning Operations Tools for Reporting Looker Analytics Unsupervised Learning

Job description

We are seeking a dynamic and highly skilled Senior Data Scientist to lead innovative data-driven initiatives across our organization. In this pivotal role, you will harness advanced analytics, machine learning, and artificial intelligence (AI) techniques to uncover insights, optimize processes, and drive strategic decision-making. Your expertise will empower teams to leverage big data technologies and cutting-edge frameworks, transforming complex datasets into actionable intelligence. The ideal candidate is passionate about exploring unstructured data, deploying scalable models, and collaborating across departments to solve challenging problems with clarity and precision., * Develop and implement sophisticated machine learning models using frameworks such as TensorFlow, Spark MLlib, and other AI tools to solve complex business problems.

  • Lead the end-to-end model training process, including data preprocessing, feature engineering, model validation, and deployment in cloud environments like AWS.
  • Design and optimize ETL pipelines for large-scale data ingestion and transformation utilizing tools such as Talend, Hadoop, Bash scripting, and SQL.
  • Collaborate with cross-functional teams to gather requirements, translate business needs into analytical solutions, and communicate insights effectively using visualization tools like Looker.
  • Conduct data mining on structured and unstructured datasets-including natural language processing (NLP) tasks-to extract valuable patterns and trends.
  • Architect scalable database designs for efficient storage and retrieval of big data using SQL, NoSQL databases, and database design best practices.
  • Stay abreast of emerging AI research such as quantum engineering applications in data science to continuously enhance analytical capabilities.

Requirements

Do you have experience in SQL?, * Extensive experience with machine learning frameworks including TensorFlow, SAS, R, Python libraries (scikit-learn, pandas), and Spark MLlib.

  • Proficiency in cloud platforms like AWS for model deployment, data storage solutions, and scalable computing resources.
  • Strong knowledge of big data technologies such as Hadoop ecosystem components (HDFS), Spark, and Talend for data integration.
  • Expertise in natural language processing (NLP), data mining techniques, unsupervised learning algorithms, and AI methodologies.
  • Solid understanding of database design principles along with hands-on experience with SQL and NoSQL databases.
  • Programming skills in Python, Java, C, Bash (Unix shell), VBA for automation and analysis tasks.
  • Familiarity with analytics tools like Looker for dashboard creation and reporting; experience with VBA scripting is a plus.
  • Ability to work on complex projects involving model training, model deployment pipelines, and continuous improvement cycles in a fast-paced environment. Join us to push the boundaries of data science innovation! Your expertise will shape the future of our analytics capabilities while working on impactful projects that challenge your skills daily.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · WWC 2024

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

3:33 min

Refactoring data science workflows using Rapids QDF and Pandas

Paul Graham Paul Graham · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:58 min

Analyzing production code coverage data using pandas

Markus Harrer Markus Harrer · WWC 2021

Videos

See all

Related articles

See all