Principal Applied Scientist, AAIS

Amazon Web Services (AWS)

Seattle (WA)

On-site

USD 199,000 - 269,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Amazon Development Center U.S., Inc. in Seattle seeks a Principal Applied Scientist to own the scientific direction of projects spanning knowledge representation, temporal reasoning, and retrieval. This is a science leadership role with hands-on prototyping and production data practices.

You will lead a team of applied scientists and MLEs, set strategy, and contribute to evaluation methods while shipping impactful results on real data.

Qualifications

  • PhD in Computer Science, ML, Statistics, or related field (or a Master’s with 8 years of applied science experience).
  • 10 years of experience building and shipping ML/AI systems reaching production users.
  • Deep expertise in large language models and at least two of information retrieval, knowledge graphs, reinforcement learning, agentic system design, or evaluation methodology for generative systems.
  • Proven experience setting technical and scientific direction for a team of scientists, including mentoring senior scientists.
  • Hands-on proficiency in Python and the ability to prototype independently in a production codebase.
  • Track record of publications, patents, or equivalent evidence of original scientific contribution.

Responsibilities

  • Set the scientific direction for knowledge representation, retrieval, and graph-based search.
  • Advance temporal reasoning and provenance to distinguish actual vs inferred data.
  • Define when AI should interrupt humans; tune precision-recall for teams.
  • Lead measurement: build evaluation across a multi-component agentic system.
  • Design synthetic data pipelines to enable testing when real data is scarce.
  • Own the learning loop from user feedback to production improvements.
  • Decide when frontier models are needed vs domain-tuned models with cost metrics.
  • Mentor scientists, review designs, publish work, and represent science externally.

Skills

Python proficiency
Production ML systems
Knowledge representation & graphs
Information retrieval
Reinforcement learning
Mentoring
Publications/patents
Team leadership

Education

PhD in Computer Science / ML / Statistics or related
Master’s + 8 years applied science experience

Tools

Python

Job description

Description AI assistants are getting genuinely good at remembering individuals: your preferences, your projects, the thread you left open last week. But that memory stops at the edge of one person’s usage. It doesn’t reach the level at which real work happens, where the knowledge that matters is spread across many people, where one person’s decision changes what everyone else should do next, and where nobody has the full picture. We’re building AI that operates at that level: a durable, accurate understanding of how a team works, used to make that team measurably faster.

We are looking for a Principal Applied Scientist to own the scientific direction of that work. This is a broad, ambiguous, high-leverage charter. The problems span knowledge representation, temporal reasoning, retrieval, agentic behavior, and the measurement science needed to know whether any of it is working. You will not be handed a well-posed problem. You will decide which problems are worth posing.

This is a science leadership role, not a solo research role. You will set direction and raise the scientific bar across a team of applied scientists and MLEs, while staying deep enough in the work to prototype an idea yourself and prove it on real data.

Key job responsibilities

Own the scientific strategy for how organizational knowledge is represented, kept current, and retrieved: extraction, entity resolution, deduplication, graph structure, and retrieval that unifies graph, semantic, keyword, and temporal search.

Advance temporal reasoning. Knowledge changes: facts are revised, decisions are reversed, priorities move. Representing what superseded what and when, and preserving the provenance to distinguish confirmed information from inferred information, is among the hardest open problems in this space.

Define the science of proactive behavior. When is it right for an AI system to interrupt a human? These are precision-critical problems where a false positive costs far more than a miss, and where the right threshold varies by team and by individual.

Lead our measurement science. Build evaluation for completeness and correctness across a multi-component agentic system, converging on a small number of trustworthy primary metrics rather than a sprawl of component scores. Judge honestly when an offline gain is real and when it is an artifact of a sparse dataset.

Build the data that doesn’t exist. The most valuable phenomena in this domain are also the rarest, which makes naturally occurring examples too scarce to learn from. Design synthetic and simulated data pipelines that generate controlled, realistic scenarios so these capabilities can be developed and tested at all.

Own the learning loop. Turn human interaction into usable training signal, and set the direction for how the system improves from explicit feedback in the near term and from passive observation over the longer term.

Make the efficiency calls. Decide where frontier models are required and where a smaller domain-tuned model is sufficient, and build the cost and capacity measurement that makes it a data-driven decision rather than an opinion.

Raise the bar across the team. Mentor scientists, review designs, publish where the work merits it, and represent the science externally to customers and to the research community.

A day in the life

You might spend the morning in a design review arguing that a proposed approach won’t survive contact with real data, the afternoon writing a prototype yourself to demonstrate the alternative, and the end of the day convincing an engineer that the capability is worth a sprint. Our sequencing is deliberate: try the idea on intuition, validate it on real data by inspection, then measure it, then operationalize it. Scientists here are expected to identify a problem, justify it, recruit others to it, and drive it into production, across whatever parts of the system that requires. Ownership follows the problem, not the org chart.

About The Team

We are a combined science, product, and engineering team building one product together. Scientists own capabilities end to end rather than individual components, because these problems don’t decompose cleanly: a single improvement typically touches extraction, storage, and retrieval at once. We invest in the tooling that makes that practical: local full-stack environments and sandboxed realistic data, so a scientist can go from idea to result in seconds rather than waiting on a deployment or on engineering support.

The work is grounded in real usage rather than benchmarks alone, which is a rare combination for science this early: real users, real data, real feedback, and a genuinely unsolved research agenda.

Basic Qualifications
  • PhD in Computer Science, Machine Learning, Statistics, or a related quantitative field; or a Master’s degree with 8 years of applied science experience
  • 10 years of experience building and shipping machine learning or AI systems that reached production users
  • Deep expertise in large language models and at least two of: information retrieval, knowledge representation and graphs, reinforcement learning, agentic system design, or evaluation methodology for generative systems
  • Demonstrated experience setting technical and scientific direction for a team of scientists, including mentoring senior scientists
  • Hands-on proficiency in Python and the ability to prototype independently in a production codebase
  • Track record of publications, patents, or equivalent evidence of original scientific contribution
Preferred Qualifications
  • Experience with agentic and multi-turn systems, including RL-based post-training, environment simulation, or agent harness evaluation
  • Experience designing evaluation frameworks for open-ended or subjective tasks where ground truth is expensive or unavailable, including synthetic data generation
  • Experience with memory, personalization, or long-horizon context systems for LLM applications
  • Experience with temporal knowledge representation, entity resolution, or knowledge graph construction at scale
  • Experience taking a product from prototype to launch under ambiguity, including making the judgment call on when quality is sufficient to ship
  • Experience with model distillation or domain-specific tuning to reduce inference cost
  • Scientific breadth across multiple ML domains, and comfort operating outside your original specialization

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.

USA, WA, Seattle – 198,900.00 – 269,000.00 USD annually

Company

Amazon Development Center U.S., Inc.

Job ID: A10553806

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Applied Scientist, AAIS
Principal Applied Scientist, AAIS

Amazon Science • Seattle (WA)

On-site
USD 199,000 - 269,000
Health insurance
RSUs
Paid time off
+1
Principal Applied Scientist, AAIS
Principal Applied Scientist, AAIS

Amazon • Seattle (WA), Northern (KY)

Hybrid
USD 199,000 - 269,000
Principal Applied Scientist, AAIS (AWS)
Principal Applied Scientist, AAIS (AWS)

Amazon • Seattle (WA)

On-site
USD 199,000 - 269,000
Sr. Applied Scientist, C360
Sr. Applied Scientist, C360

Amazon Science • Seattle (WA)

On-site
USD 167,000 - 226,000
Sr. Applied Scientist, C360
Sr. Applied Scientist, C360

Amazon • Kansas City (MO)

Hybrid
USD 167,000 - 226,000
Health insurance
401(k) matching
Paid time off
Principal Applied Scientist - AI for Life Sciences, AWS Applied AI Solutions - Life Sciences
Principal Applied Scientist - AI for Life Sciences, AWS Applied AI Solutions - Life Sciences

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 199,000 - 269,000
Health insurance
401(k) matching
Paid time off
+2
Applied Scientist, Applied AI Solutions
Applied Scientist, Applied AI Solutions

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 143,000 - 193,000
Sr. Principal Scientist, Secure Work Enablement
Sr. Principal Scientist, Secure Work Enablement

Amazon Web Services (AWS) • New York (NY)

On-site
USD 264,000 - 350,000
Health insurance
401(k) matching
Paid time off
+1
Sr. Principal Scientist, Secure Work Enablement
Sr. Principal Scientist, Secure Work Enablement

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 240,000 - 325,000
Health insurance
401(k) matching
Parental leave
+1
Applied Science Manager, GameLift
Applied Science Manager, GameLift

Amazon Web Services (AWS) • San Diego (CA)

On-site
USD 184,000 - 249,000
Health insurance
RSUs
Paid time off
+1