> Markdown version of [/jobs/ext/2610857-senior-data-scientist](https://www.wearedevelopers.com/jobs/ext/2610857-senior-data-scientist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Scientist - **Company:** Merck Sharp & Dohme LLC - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $159,600.0 - $251,200.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Data Analysis, Big Data, Bioinformatics, Computational Biology, Computer Simulation, Data Visualization, R (Programming Language), Information Retrieval, Python (Programming Language), Machine Learning, Tensorflow, Software Engineering, Data Processing, Pytorch, Deep Learning, Information Technology, Data Pipelines - **Published:** August 11, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=3d4b9e9cd5c23992 ## About the Role We invite motivated candidates from computational disciplines (for example, statistics, biostatistics, data science, computer science, bioinformatics, or related fields) as well as from natural sciences (including biochemistry, chemistry, biology, pharmacology, or related fields) to apply. The ideal candidate will have a strong foundation in data science, AI, and software engineering, and/or working fluency with assay development and pharmacology in drug discovery and an interest in working at the interface of both. The candidate will be highly motivated, creative, curious, and collaborative, with a passion for uncovering insights hidden within large data sets. This is an excellent opportunity for individuals who are intellectually curious and analytically minded. You should be able to communicate results clearly, concisely, and effectively., * Master's (with 3+ years) or Ph.D. in epidemiology, biostatistics, computational biology, bioinformatics, computer science, chemistry, biochemistry, biology, pharmacology or related field and relevant computational experience, * Strong background with computational methods (such as statistics, machine learning, AI, or software engineering) and/or in natural sciences (such as biology, chemistry, pharmacology, or related fields) * Experience analyzing data using tools such as Python or R * Possess an interest in applying computational and/or experimental scientific expertise to problems in early target and drug discovery * Collaborative, curious, and motivated to work across disciplines and build fluency outside your core area of expertise Preferred Experience and Skills: * Experience with deep learning frameworks such as TensorFlow or PyTorch * Proficient in data wrangling, mathematical modeling and data simulation, and comfortable working with large databases * Have a track record of scientific contribution, which may include publications, impactful project work, or other research achievements * Prior experience with pharmaceutical research projects supporting early target and drug discovery * Experience with multiple data modalities, e.g. high content cell/tissue imaging and/or chemical structure data, integration with other data formats and performing interpretations #EligibleforERP, Bioinformatics, Biostatistics, Computational Biology, Computational Methods, Data Modeling, Data Science, Machine Learning, Pharmaceutical Research, Python (Programming Language), R Programming, Software Engineering, Statistical Analysis ## Description We are actively recruiting a Senior Data Scientist within the Quantitative Biosciences Department at our Research Laboratories located in South San Francisco. This role will lead statistical analyses, the implementation of novel data pipelines, and develop machine learning models that enhance our ability to discover new therapeutic targets and drug candidates (molecules). This role will apply rigorous biostatistical and computational methods to assist scientists in generating novel hypotheses for lead optimization across multiple modalities and analyzing data types from in vitro assays and in vivo models. The role also includes designing and developing agentic AI assistants that enable self-service data analysis, data visualization and summarization, mathematical modeling and data simulation, rapid information retrieval and QC, synthesis of decision-critical information, and support for medical writing of regulatory documents. Working closely with bench scientists and cross-functional partners, you will provide stakeholders with an in-depth understanding of complex data while developing novel computational pipelines and methodologies., * Become part of a creative and flexible team; collaborate and work closely with scientists in biology, pharmacology, chemistry, and data science to develop strategies and novel computational approaches to detect meaningful patterns and connections in sources of biological data (e.g., biochemical, biophysical, cellular, and in vivo) * Help define and contribute to our data analysis and visualization infrastructure development, including the development and implementation of new agentic solutions * Collaborate with bench scientists, therapeutic area discovery teams and across functional teams to develop biological hypotheses for computational interrogation * Publish and present your data at conferences ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Developing an AI.SDK](https://www.wearedevelopers.com/videos/198-developing-an-ai-sdk) ## Related Articles - [The Biggest German Tech Companies](https://www.wearedevelopers.com/magazine/424-the-biggest-german-tech-companies) - [Best Companies to Work For in Germany: Top 25 Companies in 2023 ](https://www.wearedevelopers.com/magazine/33-best-companies-to-work-for-in-germany-top-25-companies-in-2023) - [Best Companies to Work For in Berlin: Top 14 Companies in 2023 ](https://www.wearedevelopers.com/magazine/188-best-companies-to-work-for-in-berlin-top-14-companies-in-2023) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Data Analyst Salary Germany](https://www.wearedevelopers.com/magazine/277-data-analyst-salary-germany)