> Markdown version of [/jobs/ext/2841557-data-scientist](https://www.wearedevelopers.com/jobs/ext/2841557-data-scientist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist - **Company:** Elsevier Inc. - **Location:** Oxford, UK - **Salary:** £65,000.0 - £105,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Content Analysis, Decision Support Systems, Graph Database, Design of User Interfaces, Information Retrieval, Python (Programming Language), Machine Learning, Natural Language Processing, Named Entity Recognition, NumPy, Tensorflow, SciPy, Search Technologies, Unstructured Data, Feature Engineering, Pytorch, Large Language Models, Deep Learning, Model Validation, Pandas, Matplotlib, Build Management, Question Answering, Scikit Learn, Information Technology, Code Testing, Data Analytics, Data Pipelines, Unsupervised Learning - **Published:** September 11, 2026 - **Apply:** https://www.adzuna.co.uk/jobs/details/5878314311 ## About the Role * Experience in data science, machine learning, artificial intelligence, NLP, statistics, applied mathematics, computer science, or a related quantitative area. * Experience working with frontier LLMs such as OpenAIs GPTs, Anthropics Claude, and Googles Gemini, including fine-tuning LLMs and/or SLMs. * Strong Python skills and a habit of writing clean, maintainable, well-tested code. * A solid grasp of machine learning fundamentals, including supervised and unsupervised learning, feature engineering, model evaluation, model selection, and performance measurement. * Experience working with structured, semi-structured, or unstructured data, especially large-scale text or content datasets. * Familiarity with common data science and machine learning tools such as Pandas, NumPy, SciPy, Scikit-learn, PyTorch, TensorFlow, or Matplotlib. * The ability to translate complex and ambiguous requirements into practical, measurable, data-driven solutions, with strong analytical thinking, problem-solving skills, and attention to quality. * Clear communication skills, a collaborative approach to working with engineering, product, and business stakeholders, and a genuine interest in building production-ready systems that deliver real user value. ## Description * Design and build machine learning, NLP, and generative AI systems for scientific discovery, knowledge extraction, decision support, and intelligent content understanding. * Work with large-scale, complex, and heterogeneous data, including scientific publications, research datasets, knowledge graphs, ontologies, taxonomies, citations, metadata, and content from every scientific discipline. * Apply the right technique to each problem, using approaches such as classification, regression, clustering, ranking, feature engineering, deep learning, embeddings, LLMs, retrieval, and generative AI. * Develop capabilities for semantic search, information retrieval, entity extraction, content classification, recommendation, ranking, summarization, question answering, and evidence-grounded generation. * Build, evaluate, fine-tune, prompt, and integrate models into robust production systems, while continuously improving quality, relevance, reliability, and user value. * Write clean, tested, production-quality Python and contribute reusable data science components, packages, and scalable data pipelines for preprocessing, inference, experimentation, monitoring, and continuous improvement. * Support deployment, monitoring, model maintenance, drift detection, automated retraining, and ongoing optimization of data science systems. * Collaborate with engineering, product, UX, analytics, research, and domain experts, and communicate technical concepts, model behavior, insights, trade-offs, and recommendations clearly to technical and non-technical audiences. Technologies: * AI * Fine-tuning * Support * Machine Learning * PyTorch * Python * TensorFlow * numpy * pandas * UX UI Design ## Related Videos - [Python Data Visualization @ Deepnote (w/ PyViz overview)](https://www.wearedevelopers.com/videos/113-python-data-visualization-deepnote-w-pyviz-overview) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) - [How to implement convenient Python bindings to C++](https://www.wearedevelopers.com/videos/618-how-to-implement-convenient-python-bindings-to-c) ## Related Articles - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)