Data Scientist

Qureight Ltd
London, UK
13 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
£52,861.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Big Data Command-Line Interface Data Cleansing Memory Management Statistical Hypothesis Testing Python (Programming Language) Logistic Regression Machine Learning Modular Design NumPy SciPy
+13 more
Data Processing Feature Engineering Model Validation Jupyter Git Pandas Matplotlib Scikit Learn Plotly Feature Selection Machine Learning Operations Software Version Control Data Pipelines

Job description

We are looking for a talented and driven individual to join our forward-thinking data science team. In this role, you will play a vital part in analysing and interpreting complex data across research projects, clinical trials, and customer studies.

You will also help shape the future of our lung disease research by developing impactful statistical models. Collaborating closely with cross-functional teams, you will have the opportunity to work alongside major pharmaceutical companies and leading research institutions worldwide.

What you will do

  • Provide statistical expertise across all initiatives, including clinical study design, clinical trial analysis and business development activities.
  • Design, expand, and maintain robust data processing pipelines, statistical analysis and machine learning workflows,and data visualisation tools for experimental and clinical research.
  • Develop and apply supervised and unsupervised machine learning approaches, including predictive modelling and clustering, to extract insights from increasingly large and complex clinical and real-world datasets.
  • Develop new and refine existing statistical analysis processes in collaboration with stakeholders and clinicians.
  • Prepare and deliver study design protocols, analysis reports, presentations, and other materials to effectively communicate insights and findings directly to clients and other stakeholders.
  • Contribute to the development of scientific materials, including abstracts, posters, conference presentations and manuscripts for publication, as required.
  • Collaborate as an innovative and creative member of a multidisciplinary team, driving novel approaches to advance the company’s mission.

Requirements

  • 2+ years of industry experience in applied data science
  • Deep understanding of probability and statistics, including power and sample size calculations, hypothesis testing,parametric and non-parametric methods, and survival analysis techniques such as Cox regression and Kaplan-Meier estimation.
  • Proven experience applying statistical methodologies and best practices-such as data preprocessing, feature engineering, and method selection-to real-world datasets, including clinical trial data, particularly within the pharmaceutical sector or in collaboration with contract research organizations.
  • Skilled in leveraging regression models-particularly linear and logistic regression-for both inference and prediction, applying techniques such as cross-validation, regularization (L1/L2), feature selection, and model evaluation using metrics including AUC, precision, recall, and calibration.
  • Experience applying supervised and unsupervised machine learning techniques to real-world datasets, including predictive modelling, classification and clustering, with an understanding of appropriate model selection, validation and evaluation.
  • Strong communication and presentation skills with the ability to communicate complex analytical findings clearly to clients, clinicians and other technical and non-technical stakeholders.
  • Ability to contribute to scientific communications, including abstracts, posters, presentations and manuscripts.
  • High level of competence with Python (Pandas, Jupyter, Scikit-learn, NumPy, SciPy, Matplotlib, Seaborn, Plotly)
  • Ability to write clean, efficient and maintainable Python code following best practices, including modular design,version control (Git), clear documentation, error handling, testing, and performance-conscious data processing (e.g.vectorization, memory management)
  • Expert skills with data wrangling and cleaning large datasets

Even better if

  • Experience with more advanced statistical models, such as mixed-effects models, is a plus
  • Ability to use matching techniques to create synthetic arms in clinical trial cohorts
  • Proficiency with the command line
  • Experience with cloud providers (AWS preferred)

Qualifications & Education

  • Degree in statistics, mathematics or a related quantitative or scientific subject

Benefits & conditions

  • A comprehensive benefits package that includes an annual bonus plan, private medical insurance, life insurance, and a contributory pension scheme

  • 25 days annual leave, plus bank holidays and enhanced maternity leave

  • A diverse work environment that brings together experts in many fields, including software engineering, devops, data science, machine learning, quality assurance, regulatory affairs, and clinical operations.

About the company

Qureight’s mission is to accelerate clinical trials and ensure breakthroughs in lung and heart disease reach patients without delay. Our AI-powered data and imaging curation platform enables the analysis of clinical imaging and other healthcare data, helping our customers bring treatments to market, faster.

We’re looking for talented people who want their work to matter. With offices in Cambridge and London, you’ll join our multidisciplinary team of clinicians, scientists, and engineers. What unites us is our open culture, continuous learning mindset, and a shared mission to help biopharma run faster, smarter trials.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.co.uk

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:54 min

Development history of scientific computation libraries and PyViz tools

Radovan Kavický · LIVE

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:33 min

Refactoring data science workflows using Rapids QDF and Pandas

Paul Graham Paul Graham · LIVE

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham · WWC 2025

1:31 min

Baseline developer skills for software data science

Markus Harrer Markus Harrer · WWC 2021

Videos

See all

Related articles

See all