Jobs · Engineering · Washington

Senior Data Scientist, Alexa For Shopping (Rufus)

Amazon · Seattle, WA · 2 wk ago
EngineeringFull-time

Key job responsibilities

  • Own the multi-agent topology (Planner → Worker → Reasoner → Loop Controller) — inter-agent communication protocols, and loop termination logic
  • Design and manage the context window strategy across agents
  • Own all system prompts, routing prompts, and chain-of-thought scaffolding across agents
  • Define what "better" means across dimensions (factual grounding, hypothesis novelty, evidence completeness, reasoning coherence) without ground-truth labels at scale
  • Own schema grounding, sparse vector indexing, and domain-scoped kNN queries
  • Own embedding strategy, intent classification accuracy, and entity extraction quality
  • Basic Qualifications
    • 4+ years of data scientist experience
    • 5+ years of data querying languages (e.g. SQL), scripting languages (e.g. Python) or statistical/mathematical software (e.g. R, SAS, Matlab, etc.) experience
    • Experience with statistical models e.g. multinomial logistic regression
    • 5+ years of working with Data & AI related technologies, including, but not limited to, AI/ML (Artificial Intelligence/Machine Learning), GenAI (Generative AI), Analytics, Database, and/or Storage experience
    • Python proficiency — statistical modeling, data manipulation (pandas, numpy, scipy), and scripting across ML pipelines and evaluation infrastructure
    • Demonstrated experience extracting structured signal from unstructured text at scale — NLP pipelines, intent classification, entity extraction, or equivalent
  • Preferred Qualifications
    • Experience with multi-agent system evaluation and independently and end-to-end Production RAG or retrieval system experience — embedding strategy, vector search, hybrid retrieval, similarity threshold calibration
    • AWS Bedrock or Strands SDK experience — or equivalent orchestration framework (LangGraph, CrewAI, AutoGen)
    • Graph database experience (Neptune, Neo4j) — schema design, traversal queries, knowledge graph construction
    • Experience scaling NLP inference pipelines — model sizing decisions, batching strategy, SageMaker or equivalent endpoint optimization
    • Business intelligence or analytics domain background — metric definitions, dimensional modeling, causal inference
    • Track record of publishing at peer-reviewed venues or presenting at industry conferences

    Qualifications

    • 4+ years of data scientist experience
    • 5+ years of data querying languages (e.g. SQL), scripting languages (e.g. Python) or statistical/mathematical software (e.g. R, SAS, Matlab, etc.) experience
    • Experience with statistical models e.g. multinomial logistic regression
    • 5+ years of working with Data & AI related technologies, including, but not limited to, AI/ML (Artificial Intelligence/Machine Learning), GenAI (Generative AI), Analytics, Database, and/or Storage experience
    • Python proficiency — statistical modeling, data manipulation (pandas, numpy, scipy), and scripting across ML pipelines and evaluation infrastructure
    • Demonstrated experience extracting structured signal from unstructured text at scale — NLP pipelines, intent classification, entity extraction, or equivalent

    Benefits

    Amazon offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave.

    Learn more about our benefits at https://amazon.jobs/en/benefits.

Similar jobs