> Markdown version of [/jobs/ext/1909163-data-scientist](https://www.wearedevelopers.com/jobs/ext/1909163-data-scientist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist - **Company:** ShanX Medtech BV - **Location:** Eindhoven, Netherlands - **Experience:** Experienced - **Salary:** €5,000.0 - €6,667.0 - **Contract:** Temporary to permanent - **Skills:** Clean Code Principles, Algorithm Design, Data Analysis, Software Applications, Big Data, Bioinformatics, Clinical Data Repository, Computer Programming, Databases, Data Cleansing, Data Structures, Data Visualization, R (Programming Language), Python (Programming Language), Machine Learning, Raw Data, Reference Data, Tensorflow, SQL Databases, Data Processing, Feature Engineering, Model Validation, Matplotlib, Scikit Learn, Information Technology, Data Analytics, Data Management, Software Library, Unsupervised Learning - **Published:** August 4, 2026 - **Apply:** https://www.adzuna.nl/details/5826812752 ## About the Role * Master's degree in data science, Computer Science, Statistics, Bioinformatics, Biomedical Engineering, or a related discipline. * 2 years' experience in a similar function * Proficiency in programming languages commonly used in data science such as Python, R, or SQL. The ability to write efficient code for data manipulation, analysis, and modeling is necessary. * Familiarity with data management and preprocessing techniques, including data cleaning, transformation, and normalization. * Machine Learning: Basic understanding of machine learning concepts and algorithms, including supervised and unsupervised learning, classification, regression, clustering, and dimensionality reduction. * Ability to create clear and informative data visualizations using tools like Matplotlib, Seaborn, or ggplot2. Proficiency in conveying complex data insights through charts, graphs, and dashboards is important. * Fluency in English. * Openness to learn about performing experimental work in a laboratory environment and with bacteria. * Transferring algorithms and communication to professional application software developers. Maintain algorithm integrity/performance. * Preferred Qualifications * Experience with machine learning libraries such as scikit-learn or TensorFlow is beneficial. * Experience with handling large-scale datasets and databases is advantageous. * Domain Knowledge: Familiarity with the fundamentals of in-vitro diagnostics, including knowledge of biomarkers, assay technologies, and clinical applications. Understanding basic microbiology and medical terminology is advantageous. * Good communication skills to non-technical stakeholders * A passion for working in a young company environment. ## Description As a (Data) Scientist at SXMT you will be responsible for the generation and development of new data analysis tools and algorithms. The role of the (Data) Scientist is pivotal in leveraging data-driven approaches to inform and accelerate diagnostic product development, ultimately leading to improved patient outcomes and healthcare delivery. This is a part-time contract for 40 hours per week, extending over three years, with the possibility of extension and potential for an increased work week. Your responsibilities include, but are not limited to: Conducting R&D: Contribution to the generation of relevant data that will be used as an input to (training of) algorithms. Data Analysis: Leading the analysis of large and complex datasets, including clinical data, and reference data generated from diagnostic assays. Applying statistical (e.g. PCA) and machine learning techniques to extract insights, identify patterns, and uncover relationships that inform diagnostic product development. Algorithm Development: Training, developing and refining algorithms and models for data analysis, interpretation, and predictive analytics. This may involve designing algorithms for diagnosis and more, risk prediction, or treatment response prediction to support diagnostic assay development but also implementation in application software. Feature Engineering: Identifying and engineering relevant features from raw data to enhance the performance and accuracy of (predictive) models. This includes preprocessing data, selecting informative features, and optimizing feature representations for improved model performance. Model Validation: Validating (predictive) models and algorithms using appropriate validation techniques, including cross-validation, bootstrapping, and holdout validation. Assessing model performance metrics such as accuracy, sensitivity, specificity, and area under the curve (AUC) to evaluate predictive performance. Data Visualization: Creating clear and informative data visualizations, including plots, charts, and graphs, to communicate results and insights effectively to stakeholders. Visualizing complex data structures and relationships to facilitate understanding and decision-making. Continuous Learning: Staying abreast of advances in data science methodologies, techniques, and tools relevant to in-vitro diagnostics. Actively participating in professional development activities, such as training programs, conferences, and workshops, to enhance skills and knowledge. ## Related Videos - [Data Governance in the Era of AI](https://www.wearedevelopers.com/videos/1622-data-governance-in-the-era-of-ai) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Kubernetes and Microservices with Multi-Model Databases](https://www.wearedevelopers.com/videos/382-kubernetes-and-microservices-with-multi-model-databases) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [Bringing Clarity to Event Streams: Enabling Analytics and AI Through Rich Metadata](https://www.wearedevelopers.com/videos/1616-bringing-clarity-to-event-streams-enabling-analytics-and-ai-through-rich-metadata) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) ## Related Articles - [Software Developer Salary in The Netherlands [2023]](https://www.wearedevelopers.com/magazine/217-software-developer-salary-in-the-netherlands-2023) - [Best Companies in the Netherlands: Top 25 Companies in 2023 ](https://www.wearedevelopers.com/magazine/193-best-companies-in-the-netherlands-top-25-companies-in-2023) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [The Netherlands – Europe’s powerhouse for software development?](https://www.wearedevelopers.com/magazine/31-the-netherlands-europe-s-powerhouse-for-software-development) - [The Biggest German Tech Companies](https://www.wearedevelopers.com/magazine/424-the-biggest-german-tech-companies)