Senior Data Scientist, Alexa For Shopping (Rufus)

Amazon.com, Inc.
Seattle, WA, United States
20 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Compensation
$159,200.0 - $215,300.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Alexa Amazon Web Services Data Analysis Communications Protocols Databases Query Languages Dimensional Modeling R (Programming Language) Graph Database Data Intelligence Python (Programming Language)
+19 more
Logistic Regression MATLAB Machine Learning Mathematical Software Neo4j Named Entity Recognition NumPy SAS (Software) SciPy Search Technologies SQL Databases System Testing Unstructured Data Data Processing Scripting Multi-Agent Systems Generative AI Pandas Machine Learning Operations

Job description

We are building a agentic intelligence system that transforms unstructured, noisy customer-data into actionable intelligence for product analytics to guide evolution of Amazon Shopping CX’s - surfacing metrics on demand and insights unprompted, without an analyst in the loop. We are solving one of the hardest problems in the agent driven data intelligence space to isolate insights from noise. This role will own multi agent system orchestration and context management; self-improving agent layer that gets measurably better over time without human intervention and reliable signal extraction from unstructured data and proactive intelligence that detects what matters before anyone asks. Our agentic system is in production. What we don’t yet have is a system that evaluates its own output quality, identifies where it fails, and closes that feedback loop automatically., As Senior Data Scientist, you will own the multi agent orchestration and the self-improvement system end-to-end. You will also own designing the overall architecture to extract insights from unstructured data at scale. You will work directly with the principal engineer, influence the technical roadmap across the team, and partner with SDE’s.

  • This role requires operating independently on problems that are not well-defined or structured, identifying and framing research challenges across broad problem areas, and delivering end-to-end solutions that have significant impact on the product.
  • Own the multi-agent topology (Planner * Worker * Reasoner * Loop Controller) - inter-agent communication protocols, and loop termination logic
  • Design and manage the context window strategy across agents
  • Own all system prompts, routing prompts, and chain-of-thought scaffolding across agents
  • Define what ā€œbetterā€ means across dimensions (factual grounding, hypothesis novelty, evidence completeness, reasoning coherence) without ground-truth labels at scale
  • Design how eval signal propagates back into prompt updates and model routing decisions
  • Own schema grounding, sparse vector indexing, and domain-scoped kNN queries
  • Own embedding strategy, intent classification accuracy, and entity extraction quality, Overview: Lynker Corporation is seeking a Senior Fisheries Data Scientist - Salmon Survival, Migration, and Mark-Recapture Modeling to support NOAA Fisheries’ Northwest Fisheries…
  • 18 days ago

Requirements

4+ years of data scientist experience

  • 5+ years of data querying languages (e.g. SQL), scripting languages (e.g. Python) or statistical/mathematical software (e.g. R, SAS, Matlab, etc.) experience
  • Experience with statistical models e.g. multinomial logistic regression
  • 5+ years of working with Data & AI related technologies, including, but not limited to, AI/ML (Artificial Intelligence/Machine Learning), GenAI (Generative AI), Analytics, Database, and/or Storage experience
  • Python proficiency - statistical modeling, data manipulation (pandas, numpy, scipy), and scripting across ML pipelines and evaluation infrastructure
  • Demonstrated experience extracting structured signal from unstructured text at scale - NLP pipelines, intent classification, entity extraction, or equivalent Preferred Qualifications

  • Experience with multi-agent system evaluation and independently and end-to-end
  • Production RAG or retrieval system experience - embedding strategy, vector search, hybrid retrieval, similarity threshold calibration
  • AWS Bedrock or Strands SDK experience - or equivalent orchestration framework (LangGraph, CrewAI, AutoGen)
  • Graph database experience (Neptune, Neo4j) - schema design, traversal queries, knowledge graph construction
  • Experience scaling NLP inference pipelines - model sizing decisions, batching strategy, SageMaker or equivalent endpoint optimization
  • Business intelligence or analytics domain background - metric definitions, dimensional modeling, causal inference
  • Track record of publishing at peer-reviewed venues or presenting at industry conferences

Benefits & conditions

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at . USA, WA, Seattle - 159,200.00 - 215,300.00 USD annually

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:54 min

Development history of scientific computation libraries and PyViz tools

Radovan Kavický · LIVE

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell Ā· LIVE

2:24 min

Comparing Neo4j and GraphQL conceptual models

William Lyon Ā· LIVE

2:18 min

Introduction to data science applications in the retail sector

Julian Joseph Ā· LIVE

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham Ā· WWC 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou Ā· Coffee With Developers

Videos

See all

Related articles

See all