> Markdown version of [/jobs/ext/2165626-senior-data-scientist](https://www.wearedevelopers.com/jobs/ext/2165626-senior-data-scientist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Scientist - **Company:** Dotdash Meredith - **Location:** Des Moines, IA, United States (Remote available) - **Experience:** Expert - **Salary:** $175,000.0 - $190,000.0 - **Contract:** Permanent contract - **Skills:** Sql Data Warehouse, A/B Testing, Artificial Intelligence, Computer Vision, Big Data, BigQuery, Software as a Service, Information Engineering, Python (Programming Language), Machine Learning, NumPy, Raw Data, Tensorflow, Standard Sql, Azure Machine Learning, Search Technologies, SQL Databases, Management of Software Versions, Pytorch, Large Language Models, Pandas, Scikit Learn, Information Technology, Production Code, Machine Learning Operations - **Published:** August 21, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/87185442/1 ## About the Role Master's degree or higher in Computer Science, Statistics, Machine Learning, Applied Mathematics, or a related quantitative field; or equivalent practical experience., You combine the modeling depth of an applied/data scientist with the pragmatism to ship end-to-end. You bring: * Strong data science fundamentals: statistics, experimental design, and evaluation methodology, with the analytical ability to turn model results into clear product and business decisions. * Demonstrated ownership of the full A/B testing lifecycle: designing experiments, running them, reading them out, and deciding; not just reporting offline metrics. * Experience designing, training, and deploying embedding models and vector retrieval (e.g., Milvus, Pinecone, or Vertex AI Vector Search) for product or content similarity at catalog scale. * Direct experience with cold-start / sparse-signal personalization: building useful recommendations from a new catalog, new users, or both. This is a core, day-one challenge of the role. * Strong Python and modern ML frameworks (PyTorch, TensorFlow, or JAX) plus the standard scientific stack (pandas, NumPy, scikit-learn). You write production-quality code, not just notebooks. * Strong SQL: hands-on experience querying large datasets in a cloud data warehouse (BigQuery preferred) to pull, join, and assemble the training and evaluation datasets that feed your models. This is a daily part of the role. * Experience deploying and serving models on a cloud ML platform: GCP Vertex AI strongly preferred (SageMaker or equivalent acceptable) and you are comfortable owning the full model lifecycle: training, deployment, versioning, and monitoring. * Commerce intuition: you've worked with product catalogs and understand merchandising, category, and PM concerns. It shows up in how you talk about catalogs and taste, not just models. * Curiosity and pragmatism about emerging AI, particularly LLMs and modern retrieval/ranking, with a track record of bringing new techniques into real production use. * Strong written and verbal communication; able to explain technical tradeoffs to both technical and non-technical stakeholders., * Applied NLP and/or computer vision for extracting structured attributes from product text and imagery. * Experience with adaptive recommendation and experimentation methods; multi-armed or contextual bandits. * Public writing or conference talks on recommendation, personalization, or ranking work. * Early-stage or commerce experience where you wore multiple hats and shipped against real business metrics (e.g., commerce SaaS or a vertical commerce startup). ## Description As a Senior Data Scientist for personalization, you will own the science behind the recommendation engine that powers each user's personalized product feed. Starting from our user-saved product signals and a live catalog ingested from thousands of retailer feeds, you will design, build, evaluate, and continuously improve the models that learn each user's taste across brand, category, color, price point, and fit. This is a hands-on, full-cycle role. You will take a recommendation problem from raw data all the way to a production model running on our existing MLOps stack - you own the model layer, not the infrastructure. Our platform team already operates the feature store, serving, and autoscaling; your job is to decide what to model, prove it works through rigorous offline and online experimentation, ship it, and iterate as behavioral signals accumulate. A defining challenge of this role is cold-start. We are launching with a small behavioral dataset and a catalog scaling from hundreds of thousands of products toward tens of millions. You will need strong commerce and product-data intuition to produce high-quality recommendations before rich click data exists - and the experimental discipline to keep improving them as it arrives. This is a foundational hire that will shape how millions of users discover products they love. Remote or Hybrid 3x a weekNYC In-office Expectations: This position offers remote work flexibility; however, if you reside within a commutable distance our offices in New York, the expectation is to work from the office three days per week., Own the design and development of the core recommendation models that turn user-saved product data into a personalized feed. Develop multi-signal models spanning brand affinity, category, color/visual attributes, fit and sizing, price sensitivity, and trend. Select and justify approaches across collaborative filtering, matrix factorization, content-based, and hybrid/neural methods (e.g., two-tower and other embedding models), and know when each applies. Build product and user embeddings that capture semantic similarity across the catalog and power candidate generation and retrieval. Design cold-start strategies that produce high-quality recommendations for new users and newly ingested products with little or no behavioral history. 25% Experimentation & Measurement Define what "good" personalization means and how it is measured. Establish rigorous offline evaluation (ranking and relevance metrics, sound holdout design) and connect it to online outcomes. Design, run, and read out A/B and multivariate experiments, and translate results into clear product and business decisions. Bring statistical discipline - sound experiment design, awareness of bias and confounding, and honest interpretation - so the team can trust which changes actually move engagement. 20% Data, Signals & Feature Understanding Develop deep intuition for Picksy's product catalog and user signals. Turn implicit behavior (saves, clicks, dwell, shares) and catalog attributes into meaningful model features, writing SQL against BigQuery to pull, join, and shape raw data into training/evaluation datasets. Apply NLP and computer-vision techniques - including modern embedding and LLM-based approaches - to extract structured attributes (category, color, material, fit) from unstructured product descriptions and imagery, and to enrich sparse catalog data. Partner with data engineering on data quality, freshness, and coverage as the catalog scales from hundreds of thousands toward tens of millions of products. 20% Full-Cycle Ownership & Productionization Take models from prototype to production yourself. Write clean, production-quality code and deploy into the existing MLOps pipeline (feature store, training, serving, monitoring) rather than building infrastructure from scratch. Own model performance in production: instrument it, watch for drift and degradation, and iterate as behavioral signals accumulate. Document models, features, and decisions clearly, and collaborate closely with the MLOps, engineering, and product teams to integrate the model layer into the live product. ## Related Videos - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) ## Related Articles - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 139 - Soft and hard queries](https://www.wearedevelopers.com/magazine/487-dev-digest-139-soft-and-hard-queries)