Sr. Data Scientist (Urology)

Johns Hopkins University
Baltimore, MD, United States
26 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$135,900.0 - $238,400.0
Working hours
Regular working hours
Languages
German
Job source

Tech stack

Artificial Neural Networks Continuous Integration Data Files Data Visualization Python (Programming Language) Machine Learning M Programming Language Software Tools Management of Software Versions Digital Twin Pytorch Deep Learning
+5 more
Git Xgboost Machine Learning Operations Software Version Control Programming Languages

Job description

We are seeking a Sr. Data Scientist. The Sr. Data Scientist will serve as a data science subject matter expert and lead the design, development, and execution of data science initiatives requiring advanced machine learning and production-level code. The Sr. Data Scientist will research and implement state-of-the-art modeling methodologies, pipelines and ML lifecycle architecture to support projects and downstream decision making. The Sr. Data Scientist will lead digital tool development efforts and conduct analytics to support full use of data procured by the Brady Urological Institute. The Sr. Data Scientist will foster collaboration efforts with data visualization specialists, data engineers, data scientists, analysts, researchers, program managers, coaches, and other external collaborators and partners in support of a portfolio of data projects., In this role, the Sr. Data Scientist will serve as the lead computational scientist for the Cancer Ecology Center within the Brady Urological Institute, owning the design, development, and production engineering of the Center’s machine-learning and simulation models. The Sr. Data Scientist will write and maintain the core modeling codebase - including next-generation development of the ExposoGraph exposome knowledge-graph platform and the build-out of the Cancer Ecology Digital Twin (CEDT), a predictive simulation environment for modeling tumor-ecosystem dynamics and individualized disease trajectories. The Sr. Data Scientist will leverage Python skills across four broad domains: classic ML (regression/classification tasks; e.g., XGBoost/LightGBM), deep learning (neural networks/ODEs; e.g., PyTorch), state-of-the-art transformer/diffusion methodologies, and causal inference (double machine learning, ATE/CATE; e.g., CausalML), as well as utilize MLOps practices such as Git versioning and CICD pipelines. These skills will facilitate translating complex, multi-modal biomedical and environmental datasets into validated, reproducible models that inform research and clinical decision-making. The successful candidate will combine deep analytical modeling and machine-learning expertise with strong software-engineering discipline, the ability to architect data and modeling pipelines from the ground up, and a track record of leading technically rigorous projects from concept to production. Experience bridging research and applied environments, fluency in complex, interpretable, and causal machine-learning methods, and the capacity to collaborate across data engineers, visualization specialists, clinicians, and research scientists are essential.

Specific Duties & Responsibilities

  • Design data modeling processes to build statistical and simulation models on complex data sets.
  • Provide detail-oriented and organized analytics and models to transform data sets into meaningful insights to inform stakeholder decision making.
  • Research best practices and state-of-the-art methodologies that can be applied in the assigned area.
  • Lead the identification of data sets for modeling which leads to quantitative conclusions.
  • Clean, assess quality and bias, explore, analyze, and visualize data; may delegate duties as necessary.
  • Document, share, and train others on methods; may delegate duties as necessary.
  • Function as a subject matter expert and/or project lead.
  • Share in-depth knowledge as a resource.
  • Train data analysts on the use of software tools to carry out analysis. Update training on a regular basis.
  • Develop models and tools related to the use of data. Provide input into other models and tools.
  • Lead cross functional teams that may include developers, analysts, data scientists, researchers, policy experts, external partners, contractors, and vendors.
  • Communicate with leadership, as well as technical and non-technical stakeholders.
  • Other duties as assigned., Johns Hopkins University requires all faculty, staff, and students to receive the seasonal flu vaccine. Exceptions to the flu vaccine requirements may be provided to individuals for religious beliefs or medical reasons. Requests for an exception must be submitted to the JHU vaccination registry.

The following additional provisions may apply, depending upon campus. Your recruiter will advise accordingly. The pre-employment physical for positions in clinical areas, laboratories, working with research subjects, or involving community contact requires documentation of immune status against Rubella (German measles), Rubeola (Measles), Mumps, Varicella (chickenpox), Hepatitis B and documentation of having received the Tdap (Tetanus, diphtheria, pertussis) vaccination. This may include documentation of having two (2) MMR vaccines; two (2) Varicella vaccines; or antibody status to these diseases from laboratory testing. Blood tests for immunities to these diseases are ordinarily included in the pre-employment physical exam except for those employees who provide results of blood tests or immunization documentation from their own health care providers. Any vaccinations required for these diseases will be given at no cost in our Occupational Health office.

Requirements

  • Master’s Degree.
  • Seven years of related experience.
  • Additional education may substitute for required experience and additional related experience may substitute for required education beyond a high school diploma/graduation equivalent, to the extent permitted by the JHU equivalency formula.

Preferred Qualifications

  • PhD.

Technical Skills & Expected Level of Proficiency

  • Data Tools and Platforms - Authority
  • Data Tool and Resource Development - Authority
  • Data Visualization - Authority
  • Machine Learning - Authority
  • Project Management - Authority
  • Programming Languages - Advanced
  • Statistical Modeling - Authority
  • Version Control System - Advanced

The core technical skills listed are most essential; additional technical skills may be required based on specific division or department needs., Please refer to the job description above to see which forms of equivalency are permitted for this position. If permitted, equivalencies will follow these guidelines: JHU Equivalency Formula: 30 undergraduate degree credits (semester hours) or 18 graduate degree credits may substitute for one year of experience. Additional related experience may substitute for required education on the same basis. For jobs where equivalency is permitted, up to two years of non-related college course work may be applied towards the total minimum education/experience required for the respective job.

Applicants Completing Studies Applicants who do not meet the posted requirements but are completing their final academic semester/quarter will be considered eligible for employment and may be asked to provide additional information confirming their academic completion date.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

1:09 min

Configuring synthetic data for safe interactive programming

Mingshen Sun Mingshen Sun · World Congress 2024

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

9:47 min

Transforming tabular metrics into meaningful business value dashboards

Boris Krumrey +2 · LIVE

Videos

See all

Related articles

See all