Senior Data Scientist I

Elsevier
Amsterdam, Netherlands
3 days ago
Apply on www.adzuna.nl
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
€4,483.0 - €7,492.0
Working hours
Shift work
Job source

Tech stack

Application Programming Interfaces (APIs) Agile Methodology Artificial Intelligence Algorithm Design Amazon Web Services Data Analysis Artificial Neural Networks JIRA Microsoft Azure Cloud Computing Data Integration Data Transformation
+27 more
Data Visualization Github Information Retrieval Python (Programming Language) Logistic Regression Machine Learning Machine Translation Natural Language Processing Software Engineering Support Vector Machine Reinforcement Learning Retrieval-Augmented Generation Transfer Learning Large Language Models Random Forest Prompt Engineering Deep Learning Generative AI Agentic-AI Gitlab Information Technology Transformer Architectures Machine Learning Operations Virtual Agents Software Version Control OpenSearch Databricks

Job description

As a Senior Data Scientist, you will play a pivotal role in the development and deployment of cutting-edge Gen AI models and solutions. You will be responsible for building, testing, and maintaining our Gen AI, RAG and NLP solutions

You will work throughout the whole life cycle of data science projects: design, implementation, production and beyond. You will deliver efficient and production-ready Python code. You will collaborate closely with developers to deploy and productionize our data science pipelines and with subject matter experts in biology and chemistry domains to validate the output.

This role requires a strong foundation in Natural Language Processing (NLP), Machine Learning, Transformer models and Generative AI, as well as proficiency in Python.

Responsibilities

  • Data collection, data analysis, model development, defining quality metrics, quality assessment of models and regular presentations to stakeholders.
  • Creating production-ready Python packages for each component of data science pipelines (such as pre-processing and model inference) and their deployment together with software engineering team
  • Optimizing and customizing Retrieval Augmented Generation (RAG) pipelines to meet specific project requirements that involve content ingestion, machine translation, and contextualized information retrieval
  • Ingesting, preprocessing, and transforming large-scale multilingual data to ensure high-quality inputs for downstream models.
  • Building AI agentic models integrated with RAG pipelines.
  • Conducting rigorous testing and evaluation of AI models to ensure high performance and reliability.
  • Integrating data science components and performing end-to-end quality assessments.
  • Maintaining robustness of data science pipelines against model drift and ensuring consistent output quality.
  • Establishing reporting processes for pipeline performance and developing automated re-training strategies for existing pipelines.
  • Collaborating with cross-functional teams to integrate AI solutions into existing products and services.
  • Leading and managing projects with a team of data scientists and independently executing the entire small-scale projects
  • Mentoring junior data scientists and fostering a knowledge-sharing culture within the team.
  • Staying up-to-date with the latest advancements in AI, machine learning, and NLP technologies.

Requirements

Are you interested in working with data and analytics to solve problems?

Are you interested in bringing your GenAI, ML and NLP expertise to projects?, * Master’s or Ph.D. in Computer Science, Data Science, Artificial Intelligence, or a related field.

  • 5+ years of relevant applied experience in data science, with a focus on Generative AI, NLP, and machine learning.
  • Proficiency in Python for data analysis, model development, and deployment.
  • Strong experience with transformer models
  • Proficiency in Generative AI technologies, including utilizing LLMs via API access, LLM evaluation tools, and prompt engineering.
  • Knowledge of various RAG pipelines and their practical implementation.
  • Experience building Agentic RAG systems is strong requirement.
  • Experience with AI agent management frameworks such as LangChain, or similar tools.
  • Experience with advanced algorithms in deep learning, neural networks, reinforcement learning, and transfer learning.
  • Familiarity with traditional machine learning algorithms such as random forests, SVM, logistic regression, and Bayesian modelling for model building, validation, and testing.
  • Familiarity with cloud platforms (e.g., Bedrock, AWS, Azure) for model deployment and the creation of production-ready pipelines.
  • Proficiency in data visualization tools and techniques.
  • Experience with version control systems (e.g., GitLab or GitHub), Jira, and working in an Agile environment.
  • Proficient in using OpenSearch and Databricks.
  • Excellent problem-solving and analytical skills, with strong attention to detail.
  • Strong communication skills and the ability to work effectively in a team-oriented environment.

Work in a way that works for you

Benefits & conditions

We promote a healthy work/life balance across the organization. We offer an appealing working prospect for our people. With numerous wellbeing initiatives, shared parental leave, study assistance and sabbaticals, we will help you meet your immediate responsibilities and your long-term goals.

  • Flexible working hours - flexing the times when you work in the day to help you fit everything in and work when you are the most productive.

About the business

As a global leader in information and analytics, we help researchers and healthcare professionals advance science and improve health outcomes for the benefit of society. Building on our publishing heritage, we combine quality information and vast data sets with analytics to support visionary science and research, health education, and interactive learning, as well as exceptional healthcare and clinical practice. At Elsevier, your work contributes to the world’s grand challenges and a more sustainable future. We harness innovative technologies to support science and healthcare to partner for a better world.

Primary Location Base Pay Range: NLD Amsterdam (Radarweg) €53,800 - €89,900.

This role is covered by the Collective Labor Agreement Publishing Industry.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.nl
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

3:14 min

Testing and environment management in GitLab CI

Martin Beránek · LIVE

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:11 min

Updating the delivery architecture with Jira and Tekton pipelines

Lian Li · World Congress 2022

Videos

See all

Related articles

See all