> Markdown version of [/jobs/ext/1418455-senior-data-scientist-i-leapspace](https://www.wearedevelopers.com/jobs/ext/1418455-senior-data-scientist-i-leapspace). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Scientist I - LeapSpace - **Company:** Elsevier B.V. - **Location:** London, UK - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** A/B Testing, Artificial Intelligence, Automated Storage and Retrieval Systems, Big Data, Computer Programming, Data Visualization, Distributed Data Store, Graph Database, Information Retrieval, Python (Programming Language), Machine Learning, Metadata Standards, Power BI, Azure Machine Learning, Search Technologies, Tableau (Software), Pytorch, Large Language Models, Matplotlib, Information Technology, HuggingFace, Machine Learning Operations, Databricks - **Published:** July 24, 2026 - **Apply:** https://relx.wd3.myworkdayjobs.com/ElsevierJobs/job/London-Wall/Senior-Data-Scientist-I---LeapSpace_R113844-1 ## About the Role This role is ideal for someone with deep hands-on experience in search/retrieval systems, RAG pipelines, and evaluation frameworks, who is ready to operate as a senior individual contributor with growing technical leadership responsibilities., * Master's or PhD in Computer Science, Data Science, Machine Learning, or a related field (or equivalent practical experience) * Experience in data science, machine learning, or applied NLP * Strong hands-on experience with: + Search and retrieval systems (lexical, vector, hybrid) + RAG pipelines and LLM-based systems + Evaluation methodologies for ML / IR / GenAI * Advanced programming skills in Python * Experience with modern ML/NLP frameworks (e.g., PyTorch, Hugging Face, LangChain, LangGraph, Haystack) * Experience working with Databricks or similar distributed data/ML platforms * Strong understanding of experimentation design and statistical analysis, * PhD in Computer Science, Data Science, Machine Learning, or a related field * Experience working with large-scale datasets (scientific, biomedical, or enterprise data) * Familiarity with scientific ontologies and metadata standards (e.g., MeSH, UMLS, ORCID, CrossRef) * Exposure to production ML systems and MLOps practices * Familiarity with data visualization and analytical tooling (e.g., Tableau, Power BI, matplotlib, seaborn, or similar) to communicate insights effectively * Experience with human-in-the-loop evaluation or annotation workflows * Publications or demonstrated applied research in IR, NLP, or generative AI ## Description Elsevier's mission is to help researchers, clinicians, and life sciences professionals advance discovery and improve health outcomes through trusted content, data, and analytics. As the landscape of science and healthcare evolves, we are pioneering intelligent discovery experiences - from Scopus AI and LeapSpace to ClinicalKey AI, PharmaPendium, and next-generation life sciences platforms. These products leverage retrieval-augmented generation (RAG), semantic search, and generative AI to make knowledge more discoverable, connected, and actionable across disciplines. The Search & AI Evaluation team sits within the Platform Data Science organization and is responsible for advancing enterprise-scale search, retrieval, and evaluation capabilities across Elsevier's global products., We are looking for a Senior Data Scientist I to lead the development and evaluation of advanced search and generative AI systems. You will own complex problem areas end-to-end, drive methodological rigor in evaluation, and contribute to the technical direction of retrieval and RAG systems., * Play a leading role in the design and optimization of lexical, vector, and hybrid retrieval systems at scale. * Help architect and improve RAG pipelines, including retrieval strategies, prompt design, and system orchestration (e.g., LangGraph-based workflows). * Help drive experimentation with embeddings, re-ranking models, and retrieval architectures to significantly improve relevance and user outcomes. * Partner with engineering to ensure robust, scalable, and production-ready implementations., * Help define and evolve evaluation strategies for search and generative AI systems across products. * Help design robust frameworks for: + IR evaluation (e.g., NDCG, recall, ranking quality) + GenAI evaluation (e.g., grounding, faithfulness, hallucination detection) * Contribute to development of evaluation datasets, gold standards, and annotation strategies. * Guide and review experimental design, including offline evaluation and A/B testing, ensuring statistical rigor and validity. * Contribute to responsible AI practices, including bias, fairness, and risk evaluation, * Apply and adapt state-of-the-art techniques in NLP, embeddings, and generative AI to production use cases. * Evaluate and integrate emerging technologies into the team's roadmap. * Contribute to knowledge graph and semantic enrichment efforts that support retrieval systems., * Collaborate with domain experts, ontology engineers, and biomedical informaticians to integrate scientific taxonomies, citation networks, and clinical ontologies into retrieval systems. * Incorporate structured data - including datasets, chemical entities, genes, drugs, clinical trials, and patient outcomes - into AI-powered discovery pipelines. * Advance Elsevier's knowledge graph and metadata integration strategy, linking research and health data for more context-aware retrieval. * Apply cutting-edge research in information retrieval, NLP, embeddings, and generative AI to continuously evolve Elsevier's discovery and evaluation stack. Collaboration & Delivery * Work closely with product, engineering, and domain experts to define and deliver impactful solutions. * Communicate findings and recommendations clearly to both technical and non-technical stakeholders. * Take ownership of projects from problem definition through experimentation and deployment. ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [REST, GraphQL, gRPC, and more: A comparison of modern API styles](https://www.wearedevelopers.com/videos/100247-rest-graphql-grpc-and-more-a-comparison-of-modern-api-styles) - [OLAP for AI Applications and why you should care](https://www.wearedevelopers.com/videos/100212-olap-for-ai-applications-and-why-you-should-care) - [Data Analytics with Microsoft Fabric: End-to-End Use Case with Data Agents](https://www.wearedevelopers.com/videos/1547-data-analytics-with-microsoft-fabric-end-to-end-use-case-with-data-agents) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [13 AI Tools You Have to Try](https://www.wearedevelopers.com/magazine/219-13-ai-tools-you-have-to-try) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud)