Lead Data scientist

Phillips, Phillip
Cambridge, MA, United States
1 day ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Working hours
Regular working hours
Job source

Tech stack

Clean Code Principles Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Business Analytics Applications Automation of Tests Microsoft Azure Code Review Continuous Integration Information Engineering Data Governance Data Security
+25 more
Decision Support Systems Python (Programming Language) Machine Learning Natural Language Processing Rapid Prototyping Process Regression Testing Tensorflow Standard Sql Search Technologies Software Deployment SQL Databases Management of Software Versions Pytorch Retrieval-Augmented Generation Large Language Models Deep Learning Model Validation Generative AI Information Technology Deployment Automation Data Analytics Machine Learning Operations GPT Software Version Control Data Pipelines

Job description

  • ML & Deep Learning Model Development
  • Design, train, and optimize ML models for prediction, classification, ranking, time-series forecasting, anomaly detection, NLP, and recommendation use cases.
  • Build robust experimentation workflows (train/validation strategy, ablations, error analysis) and improve model quality through iterative tuning.
  • Ensure reproducibility and maintainability through clean code practices, versioning, and automated testing.
  • GenAI Engineering (LLMs, RAG / MCP / fine-tuning, Agents)
  • Build enterprise-grade LLM applications using RAG (retrieval-augmented generation), MCP, and fine-tuning approaches: chunking strategies, embedding generation, hybrid retrieval, reranking, prompt templates, and citation/attribution patterns.
  • Develop LLM applications with tool use/function calling patterns and agentic workflows where appropriate.
  • Implement systematic evaluation: curated eval sets, prompt regression tests, hallucination checks, retrieval quality metrics, and automated quality gates.
  • ML & LLM Operations: Productionization, Deployment & Monitoring
  • Deploy and operate real-time and batch inference solutions on Azure using managed endpoints and/or containerized serving.
  • Build CI/CD for ML systems: automated packaging, container builds, model validation tests, staged rollouts, and rollback strategies.
  • Establish lifecycle management: model registry/versioning, lineage, promotion workflows, and release governance.
  • Implement observability: latency, throughput, cost, drift signals, data quality checks, alerts, and performance degradation monitoring.
  • Pipeline Orchestration & Automation (Train * Deploy)
  • Build standardized ML pipelines for training, evaluation, and deployment using orchestration tools (cloud-native pipelines and/or platform tools).
  • Automate dataset/version management, feature generation, scheduled retraining triggers, and approval workflows.
  • Define repeatable patterns for scalable experimentation and reliable production delivery.
  • Analytics Products, Dashboards & Data Governance
  • Own key analytics outputs as products (dashboards, reusable datasets, internal tools), continuously improving them based on usage patterns and performance gaps.
  • Build and automate dashboards and analytical components using scalable SQL logic, Python transformations, and reusable modules.
  • Act as owner for critical commercial/syndicated datasets (e.g., GfK, Circana, Nielsen or equivalent): definitions, assumptions, and limitations, ensuring transparent logic and trust in outputs.
  • Partner with data engineering/IT to ensure data quality, harmonization, and governance through strong validation and reconciliation practices.
  • Stakeholder Partnership & Decision Support (Lightweight, High Impact)
  • Serve as trusted analytics thought partner to senior stakeholders (e.g., BU leadership, Sales, Marketing, Finance), shaping problem statements and aligning on success metrics.
  • Translate complex analytics into clear recommendations with a decision-oriented storyline (“so-what / now-what”), tailored for leadership forums and reviews.
  • Support performance reviews, planning cycles, and high-priority ad-hoc requests with speed, rigor, and confidence; proactively challenge assumptions with fact-based insights.
  • Responsible AI, Security, and Risk Controls (GenAI-ready)
  • Implement guardrails: prompt injection defenses, sensitive data protections, output validation, and secure tool execution patterns.
  • Apply responsible AI practices: transparent evaluation criteria, auditability, and risk controls aligned to enterprise needs.
  • Technical Leadership (Lead-level Expectations)
  • Set engineering standards for DS/ML codebases: design docs, code review practices, testing discipline, and production readiness checklists.
  • Mentor data scientists/ML engineers on modeling, GenAI engineering, and MLOps best practices.
  • Lead architectural decisions across modeling approaches, retrieval stack, serving patterns, and evaluation strategy.

Requirements

  • Strong Python (production-quality coding) and solid CS fundamentals; strong SQL for data access and validation.
  • Depth in ML: Traditional ML exposure and at least one deep learning framework (PyTorch/TensorFlow), with strong understanding of metrics and failure modes.
  • GenAI implementation: RAG / MCP / fine-tuning, embeddings/vector search, prompt orchestration, evaluation harnesses, and LLM application patterns.
  • Production deployment experience on AWS or Azure (model/LLM app deployment, API serving, scaling, monitoring).
  • MLOps tooling: experiment tracking, model registry, CI/CD, and pipeline orchestration (e.g., MLflow or equivalent patterns).

Good-to-have (Business + Influence)

  • Strong business acumen and ability to connect disparate data points into compelling narratives that influence senior stakeholders.
  • Builder/MVP mindset-rapid prototyping and iterating based on stakeholder feedback while maintaining data quality and governance, * Bachelor’s degree in engineering, Computer Science, Statistics, Economics, Mathematics, or a related quantitative field.
  • Master’s degree preferred (e.g., Data Analytics, Business Analytics, Applied Statistics, Economics, AI, or MBA with strong analytics focus).
  • Continuous learning mindset expected, with demonstrated upskilling in advanced analytics, AI, or data engineering concepts (formal or informal).

Note: This role values applied problem-solving and business impact over purely academic specialization.

You’re The Right Fit If

  • Proven track record of owning end-to-end analytics domains, not just contributing to isolated analyses or consuming pre-built reports.
  • 7-12+ years in hands-on Data Science / ML Engineering with multiple production deployments owned end-to-end.
  • Demonstrated ability to take solutions from experimentation * production (reproducible pipelines, deployment to managed endpoints/container platforms, monitoring + iterative improvement).
  • Strong GenAI delivery record: shipped RAG/MCP/fine-tuned LLM applications with measurable quality controls, safety measures, and operational readiness.
  • Experience operating in complex, matrixed environments and partnering with senior stakeholders to drive insight-led decision making
  • Hands-on exposure to AI-enabled analytics, including the use of GenAI tools (e.g., ChatGPT, Claude, or similar) to accelerate insight generation, analysis, or productivity.
  • Strong experience partnering with senior business stakeholders (BU leaders, Sales, Marketing, Finance), influencing decisions through insight-led storytelling.

About the company

About Philips

We are a health technology company. We built our entire company around the belief that every human matters, and we won’t stop until everybody everywhere has access to the quality healthcare that we all deserve. Do the work of your life to help the lives of others.

  • Learn more about our business.
  • Discover our rich and exciting history.
  • Learn more about our purpose.

If you’re interested in this role and have many, but not all, of the experiences needed, we encourage you to apply. You may still be the right candidate for this or other opportunities at Philips. Learn more about our culture of impact with care here., We don’t want you to be satisfied with your job. We want you to be curious, inspired, challenged, and most of all, excited about new possibilities. At Philips, what you do every day can contribute to innovative health technologies and solutions that make a positive and very visible impact on billions of people every year. Including you…

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · World Congress 2024

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

1:39 min

Fundamentals of tensors and the TensorFlow library

Håkan Silfvernagel · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:53 min

Architecting machine learning projects with the PAI platform

Qiyang Duan · LIVE

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · World Congress 2025

Videos

See all

Related articles

See all