Principal Applied Scientist, AAIS (AWS)

Amazon

Seattle (WA)

On-site

USD 199,000 - 269,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Amazon Development Center U.S., Inc. seeks an experienced scientist to own the scientific strategy for organizational knowledge representation, retrieval, and graph integration.

You will lead efforts in extraction, entity resolution, and data provenance while shaping proactive AI behavior and evaluation With a background in large language models and multiple ML domains, you will mentor junior scientists, design scalable evaluation, and determine when to deploy frontier vs.

Qualifications

  • PhD in CS/ML/Statistics or related field, or Master’s with 8+ years applied science experience.
  • 10+ years of experience building and shipping ML/AI systems to production users.
  • Deep expertise in large language models and at least two of: information retrieval, knowledge representation and graphs, reinforcement learning, agentic system design, or evaluation methodology for generative systems.
  • Demonstrated experience setting technical and scientific direction for a team of scientists, including mentoring senior scientists.
  • Hands-on proficiency in Python and the ability to prototype independently in a production codebase.
  • Track record of publications, patents, or equivalent evidence of original scientific contribution.

Responsibilities

  • Own the scientific strategy for how organizational knowledge is represented, kept current, and retrieved: extraction, entity resolution, deduplication, graph structure, and retrieval that unifies graph, semantic, keyword, and temporal search.
  • Advance temporal reasoning. Knowledge changes: facts are revised, decisions are reversed, priorities move. Representing what superseded what and when, and preserving the provenance to distinguish confirmed information from inferred information, is among the hardest open problems in this space.
  • Define the science of proactive behavior. When is it right for an AI system to interrupt a human? These are precision‑critical problems where a false positive costs far more than a miss, and where the right threshold varies by team and by individual.
  • Lead our measurement science. Build evaluation for completeness and correctness across a multi‑component agentic system, converging on a small number of trustworthy primary metrics rather than a sprawl of component scores. Judge honestly when an offline gain is real and when it is an artifact of a sparse dataset.
  • Build the data that doesn't exist. The most valuable phenomena in this domain are also the rarest, which makes naturally occurring examples too scarce to learn from. Design synthetic and simulated data pipelines that generate controlled, realistic scenarios so these capabilities can be developed and tested at all.
  • Own the learning loop. Turn human interaction into usable training signal, and set the direction for how the system improves from explicit feedback in the near term and from passive observation over the longer term.
  • Make the efficiency calls. Decide where frontier models are required and where a smaller domain‑tuned model is sufficient, and build the cost and capacity measurement that makes it a data‑driven decision rather than an opinion.
  • Raise the bar across the team. Mentor scientists, review designs, publish where the work merits it, and represent the science externally to customers and to the research community.

Skills

Python programming
Mentoring senior scientists
Publications or patents

Education

PhD in Computer Science / ML / Statistics or related field
Master's degree with 8+ years applied science experience

Tools

Python

Job description

Job ID: 10553806 | Amazon Development Center U.S., Inc.

AI assistants are getting genuinely good at remembering individuals: your preferences, your projects, the thread you left open last week. But that memory stops at the edge of one person's usage. It doesn't reach the level at which real work happens, where the knowledge that matters is spread across many people, where one person's decision changes what everyone else should do next, and where nobody has the full picture. We're building AI that operates at that level: a durable, accurate understanding of how a team works, used to make that team measurably faster.

Key job responsibilities
  • Own the scientific strategy for how organizational knowledge is represented, kept current, and retrieved: extraction, entity resolution, deduplication, graph structure, and retrieval that unifies graph, semantic, keyword, and temporal search.
  • Advance temporal reasoning. Knowledge changes: facts are revised, decisions are reversed, priorities move. Representing what superseded what and when, and preserving the provenance to distinguish confirmed information from inferred information, is among the hardest open problems in this space.
  • Define the science of proactive behavior. When is it right for an AI system to interrupt a human? These are precision‑critical problems where a false positive costs far more than a miss, and where the right threshold varies by team and by individual.
  • Lead our measurement science. Build evaluation for completeness and correctness across a multi‑component agentic system, converging on a small number of trustworthy primary metrics rather than a sprawl of component scores. Judge honestly when an offline gain is real and when it is an artifact of a sparse dataset.
  • Build the data that doesn't exist. The most valuable phenomena in this domain are also the rarest, which makes naturally occurring examples too scarce to learn from. Design synthetic and simulated data pipelines that generate controlled, realistic scenarios so these capabilities can be developed and tested at all.
  • Own the learning loop. Turn human interaction into usable training signal, and set the direction for how the system improves from explicit feedback in the near term and from passive observation over the longer term.
  • Make the efficiency calls. Decide where frontier models are required and where a smaller domain‑tuned model is sufficient, and build the cost and capacity measurement that makes it a data‑driven decision rather than an opinion.
  • Raise the bar across the team. Mentor scientists, review designs, publish where the work merits it, and represent the science externally to customers and to the research community.

A day in the life
You might spend the morning in a design review arguing that a proposed approach won't survive contact with real data, the afternoon writing a prototype yourself to demonstrate the alternative, and the end of the day convincing an engineer that the capability is worth a sprint. Our sequencing is deliberate: try the idea on intuition, validate it on real data by inspection, then measure it, then operationalize it. Scientists here are expected to identify a problem, justify it, recruit others to it, and drive it into production, across whatever parts of the system that requires. Ownership follows the problem, not the org chart.

About the team
We are a combined science, product, and engineering team building one product together. Scientists own capabilities end to end rather than individual components, because these problems don't decompose cleanly: a single improvement typically touches extraction, storage, and retrieval at once. We invest in the tooling that makes that practical: local full‑stack environments and sandboxed realistic data, so a scientist can go from idea to result in seconds rather than waiting on a deployment or on engineering support.

The work is grounded in real usage rather than benchmarks alone, which is a rare combination for science this early: real users, real data, real feedback, and a genuinely unsolved research agenda.

Basic Qualifications
  • PhD in Computer Science, Machine Learning, Statistics, or a related quantitative field; or a Master's degree with 8+ years of applied science experience
  • 10+ years of experience building and shipping machine learning or AI systems that reached production users
  • Deep expertise in large language models and at least two of: information retrieval, knowledge representation and graphs, reinforcement learning, agentic system design, or evaluation methodology for generative systems
  • Demonstrated experience setting technical and scientific direction for a team of scientists, including mentoring senior scientists
  • Hands‑on proficiency in Python and the ability to prototype independently in a production codebase
  • Track record of publications, patents, or equivalent evidence of original scientific contribution
Preferred Qualifications
  • Experience with agentic and multi‑turn systems, including RL‑based post‑training, environment simulation, or agent harness evaluation
  • Experience designing evaluation frameworks for open‑ended or subjective tasks where ground truth is expensive or unavailable, including synthetic data generation
  • Experience with memory, personalization, or long‑horizon context systems for LLM applications
  • Experience with temporal knowledge representation, entity resolution, or knowledge graph construction at scale
  • Experience taking a product from prototype to launch under ambiguity, including making the judgment call on when quality is sufficient to ship
  • Experience with model distillation or domain‑specific tuning to reduce inference cost
  • Scientific breadth across multiple ML domains, and comfort operating outside your original specialization

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign‑on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .

USA, WA, Seattle - 198,900.00 - 269,000.00 USD annually

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Applied Scientist, AAIS
Principal Applied Scientist, AAIS

Amazon Science • Seattle (WA)

On-site
USD 199,000 - 269,000
Health insurance
RSUs
Paid time off
+1
Principal Applied Scientist, AAIS
Principal Applied Scientist, AAIS

Amazon • Seattle (WA), Northern (KY)

Hybrid
USD 199,000 - 269,000
Principal Applied Scientist, AAIS
Principal Applied Scientist, AAIS

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 199,000 - 269,000
Sr. Applied Scientist, C360
Sr. Applied Scientist, C360

Amazon Science • Seattle (WA)

On-site
USD 167,000 - 226,000
Senior Applied Scientist, Kumo
Senior Applied Scientist, Kumo

Amazon • Bellevue (WA)

Hybrid
USD 167,000 - 226,000
Health insurance
401(k) matching
Paid time off
+1
Senior Applied Scientist, Kumo
Senior Applied Scientist, Kumo

Amazon • Factoria (WA)

Hybrid
USD 167,000 - 226,000
Senior Applied Scientist, Kumo
Senior Applied Scientist, Kumo

Socket.dev • Bellevue (WA)

Hybrid
USD 167,000 - 226,000
RSUs
Health insurance
401(k) matching
+1
Senior Applied Scientist, Kumo
Senior Applied Scientist, Kumo

Amazon Science • Bellevue (WA)

Hybrid
USD 167,000 - 226,000
Applied Scientist, AWS Quick
Applied Scientist, AWS Quick

Amazon Science • New York (NY)

On-site
USD 172,000 - 223,000
Health insurance
401(k) matching
Paid time off
+1
Applied Scientist, AWS Quick
Applied Scientist, AWS Quick

Amazon Science • Santa Clara (CA)

On-site
USD 172,000 - 222,000
Health insurance
RSUs
401(k) matching
+1