AI Session Reviewer - English Expert - AI Trainer

Mercor

Seattle (WA)

On-site

USD 28,000 - 39,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Mercor is hiring English-speaking reviewers to evaluate recorded sessions with a voice-and-camera AI assistant. You will watch each session, compare what the assistant said with what was visible, and score it against a detailed rubric.

Your written comments are the deliverable, and you will flag failures for correction. This is judgment work, not volume work, with careful attention to safety and accuracy in every review.

Qualifications

  • Native English proficiency, both spoken and written.
  • Clear, precise written English; the scoring guide is written in English.
  • An Android phone with a working camera and microphone is mandatory for the project.
  • Attention to detail and the discipline to apply a rubric consistently.

Responsibilities

  • Watch sessions and compare assistant dialogue to video visuals.
  • Score each session on interpretation accuracy, relevance, and timing.
  • Identify hallucinations or fabricated visual details.
  • Assess whether safety warnings are timely and proportionate.
  • Provide detailed turn-by-turn comments in the intake form.

Skills

Native English
Attention to detail
Discipline to apply rubric
Written English proficiency

Tools

Android phone

Job description

About the Role

Mercor is hiring English-speaking reviewers to evaluate recorded sessions with a voice-and-camera AI assistant. You will watch each session, compare what the assistant said against what was actually visible in the video, and score it against a detailed rubric.

Your written comments are the deliverable. What you flag as a failure is what gets fixed first. This is judgment work, not volume work.

Key Responsibilities
  • Review recorded sessions in which a user speaks to an AI assistant while their camera streams whatever is in front of them.

  • Compare the assistant's responses against what was actually visible in the video feed, and identify hallucinated or fabricated visual detail.

  • Score each session across visual interpretation, accuracy, relevance, response timing, and whether the assistant used spatial and directional language ("to your left", "move your finger up") rather than purely visual description.

  • Judge whether the assistant appropriately warned the user when its answer could affect their health, physical safety, or financial security, and whether that warning was timely and proportionate.

  • Write detailed open comments that support every score with concrete evidence from the session.

  • Capture the conversation turn by turn in the intake form.

Required Qualifications
  • Native or native-level English, spoken and written.

  • Clear, precise written English; the scoring guide is written in English.

  • An Android phone with a working camera and microphone. This is mandatory for the project.

  • Attention to detail and the discipline to apply a rubric consistently, without substituting your own criteria along the way.

Preferred Qualifications
  • Prior data annotation, linguistic QA, content moderation, or AI model evaluation experience.

  • Training in linguistics, translation or interpreting,

  • An ear for regional register: recognizing when the assistant's English sounds unnatural or non-native.

Additional Information
  • Start date: Immediate. High volume of work between 26 to 31 Aug is expected.

  • Content note: some sessions involve situations where a wrong answer would carry real consequences, such as identifying medication, checking whether food has spoiled, crossing a street, or reading bank card details. There is no graphic or violent content, but you will evaluate cases with safety implications, and part of the work is flagging when the assistant failed to warn about the risk.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Session Reviewer - English Expert
AI Session Reviewer - English Expert

Mercor • San Francisco (CA)

On-site
USD 45,000 - 70,000
AI Session Evaluator & Training Specialist
AI Session Evaluator & Training Specialist

Mercor • Seattle (WA)

On-site
USD 28,000 - 39,000
AI Session Quality Auditor
AI Session Quality Auditor

Mercor • San Francisco (CA)

On-site
USD 45,000 - 70,000
AI Assistant Output Evaluation Specialist
AI Assistant Output Evaluation Specialist

OpenTrain AI • United States

Remote
GBP 31,000 - 92,000
Remote work
English Language Expert
English Language Expert

SME Careers • Wichita (KS)

On-site
USD 34,000 - 55,000
English Language Expert
English Language Expert

SME Careers • Portland (ME)

On-site
USD 34,000 - 55,000
Korean Linguistic QA Specialist — AI Video Review
Korean Linguistic QA Specialist — AI Video Review

Obsidian • San Francisco (CA)

Remote
USD 60,000 - 75,000
English Language Expert
English Language Expert

SME Careers • Bloomington (IN)

On-site
USD 28,000 - 50,000
Spoken Instruction Conversation Evaluator
Spoken Instruction Conversation Evaluator

AuraOne • United States

On-site
USD 27,552,000 - 49,593,600
AI Evaluation & Annotation Reviewer (L3 - Advanced Level) - Italian (US)
AI Evaluation & Annotation Reviewer (L3 - Advanced Level) - Italian (US)

Volga Partners • Northern (KY)

Hybrid
USD 19,000 - 25,000