Data Scientist

Bayer SAS
Barcelona, Spain
4 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
2 years minimum
Working hours
Regular working hours
Languages
English

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Continuous Integration Data Governance Python (Programming Language) Machine Learning NumPy Tensorflow Standard Sql Management of Software Versions
+12 more
Data Logging Pytorch Large Language Models Multi-Agent Systems Prompt Engineering Apache Spark Model Validation Generative AI Pandas Git Flow Scikit Learn Databricks

Job description

Experteer Overview Join Bayer’s Machine Learning u**amp; AI unit to turn complex business questions into measurable AI-driven impact across functions like Finance, Supply Chain, HR, and more.You will combine GenAI with classical ML, lead rigorous experimentation, and deliver production-ready AI solutions on a cloud-native stack.Collaborate with cross-functional partners to drive adoption and scale responsible AI in a global setting.This role offers a chance to shape enterprise AI initiatives and contribute to Bayer’s mission of empowering health and sustainability.Compensaciones / Beneficios* Translate business needs into DS problems with clear hypotheses and success metrics* Design, build, and evaluate GenAI solutions (LLMs, agent workflows, embeddings) and classical ML models* Establish rigorous evaluation plans including offline metrics and human-in-the-loop reviews* Develop production-grade Python code with Git workflows, tests, docs, and reproducibility* Monitor data/feature drift, model/prompt versioning, and cost/latency; implement structured logging* Collaborate with Product, Data Engineers, AI Engineers, and stakeholders to drive adoption* Lead workshops and stakeholder storytelling on insights, risks, and trade-offsResponsabilidades* Master’s or PhD with 2+ years in Data Science or Applied ML* Strong Python and SQL skills; pandas, NumPy, scikit-learn; PyTorch or TensorFlow a plus* Hands-on with Generative AI: embeddings, prompt engineering, tool calls; agent frameworks (LangChain, LangGraph, PydanticAI)* Experience with vector databases (pgvector)* Solid ML fundamentals: model selection, validation, metrics; time series/forecasting a plus* Experience with offline/online tests, GenAI evaluation, safety checks, LangSmith/Langfuse beneficial* Clear stakeholder communication and ability to influence decisions with data* Basic cloud proficiency (AWS/Azure): storage/compute, Databricks or Spark; CI/CD familiarity* Good engineering hygiene: modular code, testing, documentation, reproducibility* Data governance and privacy awareness* Fluent in English; additional languages a plusRequisitos principales*

Requirements

  • Master’s or PhD with 2+ years in Data Science or Applied ML
  • Strong Python and SQL skills; pandas, NumPy, scikit-learn; PyTorch or TensorFlow a plus
  • Hands-on with Generative AI: embeddings, prompt engineering, tool calls; agent frameworks (LangChain, LangGraph, PydanticAI)
  • Experience with vector databases (pgvector)
  • Solid ML fundamentals: model selection, validation, metrics; time series/forecasting a plus
  • Experience with offline/online tests, GenAI evaluation, safety checks, LangSmith/Langfuse beneficial
  • Clear stakeholder communication and ability to influence decisions with data
  • Basic cloud proficiency (AWS/Azure): storage/compute, Databricks or Spark; CI/CD familiarity
  • Good engineering hygiene: modular code, testing, documentation, reproducibility
  • Data governance and privacy awareness
  • Fluent in English; additional languages a plus

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · WWC 2024

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham · WWC 2025

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

Videos

See all

Related articles

See all