> Markdown version of [/jobs/ext/3023150-lead-data-scientist](https://www.wearedevelopers.com/jobs/ext/3023150-lead-data-scientist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Data scientist - **Company:** Phillips, Phillip - **Location:** Cambridge, MA, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Clean Code Principles, Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, Business Analytics Applications, Automation of Tests, Microsoft Azure, Code Review, Continuous Integration, Information Engineering, Data Governance, Data Security, Decision Support Systems, Python (Programming Language), Machine Learning, Natural Language Processing, Rapid Prototyping Process, Regression Testing, Tensorflow, Standard Sql, Search Technologies, Software Deployment, SQL Databases, Management of Software Versions, Pytorch, Retrieval-Augmented Generation, Large Language Models, Deep Learning, Model Validation, Generative AI, Information Technology, Deployment Automation, Data Analytics, Machine Learning Operations, GPT, Software Version Control, Data Pipelines - **Published:** September 21, 2026 - **Apply:** https://arc.dev/remote-jobs/j/redirect/plrf52ni8c ## About the Role * Strong Python (production-quality coding) and solid CS fundamentals; strong SQL for data access and validation. * Depth in ML: Traditional ML exposure and at least one deep learning framework (PyTorch/TensorFlow), with strong understanding of metrics and failure modes. * GenAI implementation: RAG / MCP / fine-tuning, embeddings/vector search, prompt orchestration, evaluation harnesses, and LLM application patterns. * Production deployment experience on AWS or Azure (model/LLM app deployment, API serving, scaling, monitoring). * MLOps tooling: experiment tracking, model registry, CI/CD, and pipeline orchestration (e.g., MLflow or equivalent patterns). Good-to-have (Business + Influence) * Strong business acumen and ability to connect disparate data points into compelling narratives that influence senior stakeholders. * Builder/MVP mindset-rapid prototyping and iterating based on stakeholder feedback while maintaining data quality and governance, * Bachelor's degree in engineering, Computer Science, Statistics, Economics, Mathematics, or a related quantitative field. * Master's degree preferred (e.g., Data Analytics, Business Analytics, Applied Statistics, Economics, AI, or MBA with strong analytics focus). * Continuous learning mindset expected, with demonstrated upskilling in advanced analytics, AI, or data engineering concepts (formal or informal). Note: This role values applied problem-solving and business impact over purely academic specialization. You're The Right Fit If * Proven track record of owning end-to-end analytics domains, not just contributing to isolated analyses or consuming pre-built reports. * 7-12+ years in hands-on Data Science / ML Engineering with multiple production deployments owned end-to-end. * Demonstrated ability to take solutions from experimentation * production (reproducible pipelines, deployment to managed endpoints/container platforms, monitoring + iterative improvement). * Strong GenAI delivery record: shipped RAG/MCP/fine-tuned LLM applications with measurable quality controls, safety measures, and operational readiness. * Experience operating in complex, matrixed environments and partnering with senior stakeholders to drive insight-led decision making * Hands-on exposure to AI-enabled analytics, including the use of GenAI tools (e.g., ChatGPT, Claude, or similar) to accelerate insight generation, analysis, or productivity. * Strong experience partnering with senior business stakeholders (BU leaders, Sales, Marketing, Finance), influencing decisions through insight-led storytelling. ## Description * ML & Deep Learning Model Development * Design, train, and optimize ML models for prediction, classification, ranking, time-series forecasting, anomaly detection, NLP, and recommendation use cases. * Build robust experimentation workflows (train/validation strategy, ablations, error analysis) and improve model quality through iterative tuning. * Ensure reproducibility and maintainability through clean code practices, versioning, and automated testing. * GenAI Engineering (LLMs, RAG / MCP / fine-tuning, Agents) * Build enterprise-grade LLM applications using RAG (retrieval-augmented generation), MCP, and fine-tuning approaches: chunking strategies, embedding generation, hybrid retrieval, reranking, prompt templates, and citation/attribution patterns. * Develop LLM applications with tool use/function calling patterns and agentic workflows where appropriate. * Implement systematic evaluation: curated eval sets, prompt regression tests, hallucination checks, retrieval quality metrics, and automated quality gates. * ML & LLM Operations: Productionization, Deployment & Monitoring * Deploy and operate real-time and batch inference solutions on Azure using managed endpoints and/or containerized serving. * Build CI/CD for ML systems: automated packaging, container builds, model validation tests, staged rollouts, and rollback strategies. * Establish lifecycle management: model registry/versioning, lineage, promotion workflows, and release governance. * Implement observability: latency, throughput, cost, drift signals, data quality checks, alerts, and performance degradation monitoring. * Pipeline Orchestration & Automation (Train * Deploy) * Build standardized ML pipelines for training, evaluation, and deployment using orchestration tools (cloud-native pipelines and/or platform tools). * Automate dataset/version management, feature generation, scheduled retraining triggers, and approval workflows. * Define repeatable patterns for scalable experimentation and reliable production delivery. * Analytics Products, Dashboards & Data Governance * Own key analytics outputs as products (dashboards, reusable datasets, internal tools), continuously improving them based on usage patterns and performance gaps. * Build and automate dashboards and analytical components using scalable SQL logic, Python transformations, and reusable modules. * Act as owner for critical commercial/syndicated datasets (e.g., GfK, Circana, Nielsen or equivalent): definitions, assumptions, and limitations, ensuring transparent logic and trust in outputs. * Partner with data engineering/IT to ensure data quality, harmonization, and governance through strong validation and reconciliation practices. * Stakeholder Partnership & Decision Support (Lightweight, High Impact) * Serve as trusted analytics thought partner to senior stakeholders (e.g., BU leadership, Sales, Marketing, Finance), shaping problem statements and aligning on success metrics. * Translate complex analytics into clear recommendations with a decision-oriented storyline ("so-what / now-what"), tailored for leadership forums and reviews. * Support performance reviews, planning cycles, and high-priority ad-hoc requests with speed, rigor, and confidence; proactively challenge assumptions with fact-based insights. * Responsible AI, Security, and Risk Controls (GenAI-ready) * Implement guardrails: prompt injection defenses, sensitive data protections, output validation, and secure tool execution patterns. * Apply responsible AI practices: transparent evaluation criteria, auditability, and risk controls aligned to enterprise needs. * Technical Leadership (Lead-level Expectations) * Set engineering standards for DS/ML codebases: design docs, code review practices, testing discipline, and production readiness checklists. * Mentor data scientists/ML engineers on modeling, GenAI engineering, and MLOps best practices. * Lead architectural decisions across modeling approaches, retrieval stack, serving patterns, and evaluation strategy. ## Related Videos - [Beyond GPT: Building Unified GenAI Platforms for the Enterprise of Tomorrow](https://www.wearedevelopers.com/videos/1525-beyond-gpt-building-unified-genai-platforms-for-the-enterprise-of-tomorrow) - [ Evaluating AI models for code comprehension](https://www.wearedevelopers.com/videos/1462-evaluating-ai-models-for-code-comprehension) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Developer Experience, Platform Engineering and AI powered Apps](https://www.wearedevelopers.com/videos/990-developer-experience-platform-engineering-and-ai-powered-apps) - [Speak, Code, Deploy: Transforming Developer Experience with Voice Commands](https://www.wearedevelopers.com/videos/1159-speak-code-deploy-transforming-developer-experience-with-voice-commands) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)