Health AI Conversation Evaluator

Mercor

Greater London

Remote

GBP 21,000 - 166,000

Full time

13 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Mercor is seeking a Generalist Annotator for Health AI Conversation Quality Evaluation on a contract basis. You will evaluate AI responses, detect sycophancy, and assess tone to ensure users leave conversations informed and calm.

You will work independently, reviewing conversations from 1 to 7 turns, spending 2–13 minutes per conversation, and provide concise written justifications and improvement suggestions in clear English.

Qualifications

  • Strong written English is required.
  • Ability to explain judgments in two to four sentences.
  • Careful reading to notice omissions in responses.
  • Comfort with applying a detailed rubric consistently.
  • Reliable availability for time-boxed batches.

Responsibilities

  • Evaluate the relevance and clarity of AI responses. Ensure the AI answers the question in understandable language for non-experts.
  • Detect sycophancy in AI interactions. Identify when the AI tells users what they want to hear instead of what they need.
  • Assess the tone and outcome of conversations. Determine if the user leaves informed and calm or anxious and uninformed.
  • Provide written justification for ratings below the top option. Explain what went wrong and suggest improvements.
  • Work independently to review conversations ranging from 1 to 7 turns. Spend 2–13 minutes per conversation based on length.

Skills

Strong written English
Explain judgments
Attention to detail
Rubric-based evaluation
Time management

Job description

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.

Position: Generalist Annotator — Health AI Conversation Quality Evaluation
Type: Contract
Compensation: $20–$160/hour
Location: Remote

Role Responsibilities
  • Evaluate the relevance and clarity of AI responses. Ensure the AI answers the question in understandable language for non-experts.
  • Detect sycophancy in AI interactions. Identify when the AI tells users what they want to hear instead of what they need.
  • Assess the tone and outcome of conversations. Determine if the user leaves informed and calm or anxious and uninformed.
  • Provide written justification for ratings below the top option. Explain what went wrong and suggest improvements.
  • Work independently to review conversations ranging from 1 to 7 turns. Spend 2–13 minutes per conversation based on length.
Qualifications
Must-Have
  • Strong written English skills.
  • Ability to explain judgments in two to four specific sentences.
  • Careful reading to notice omissions in responses.
  • Comfort with applying a detailed rubric consistently.
  • Reliable availability for time-boxed batches.
Preferred
  • Experience in annotation, human-feedback, evaluation, or content review.
  • Background in explaining technical concepts to non-experts.
  • Experience with rubric-based grading or quality assurance.
Important
  • Using an AI tool to write comments is prohibited. This work captures human judgment that AI lacks. AI-written feedback corrupts the dataset and is checked for.
Resources & Support
  • For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome
  • For any help or support, reach out to: support@mercor.com

PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Health AI Conversation Quality Evaluator (Contract)
Remote Health AI Conversation Quality Evaluator (Contract)

Mercor • Greater London

Remote
GBP 21,000 - 166,000
Remote Health Communications Evaluator: AI Quality Feedback
Remote Health Communications Evaluator: AI Quality Feedback

Mercor • Greater London

Remote
GBP 34,000 - 55,000
Remote work
UI Designer - Fully Remote
UI Designer - Fully Remote

Mercor • Greater London

Remote
GBP 52,000 - 83,000
Remote Clinical AI Quality Evaluator
Remote Clinical AI Quality Evaluator

Mercor • Greater London

Remote
GBP 28,000 - 55,000
Stem Data Specialist - Engineering
Stem Data Specialist - Engineering

Mercor • England

Remote
GBP 21,000 - 83,000
AI Session Reviewer - English Expert
AI Session Reviewer - English Expert

Mercor • Greater London

On-site
GBP 24,000 - 29,000
Remote AI Quality Evaluator – Software/Data Artifacts
Remote AI Quality Evaluator – Software/Data Artifacts

Mercor • Greater London

Remote
GBP 34,000 - 55,000
Data Annotator (SaaS) | $37/hr Remote | Mercor
Data Annotator (SaaS) | $37/hr Remote | Mercor

Crossing Hurdles • United Kingdom

On-site
GBP 31,233 - 38,520
Remote Healthcare AI Quality Evaluator — Expert Reviewer
Remote Healthcare AI Quality Evaluator — Expert Reviewer

Mercor • Greater London

On-site
GBP 28,000 - 55,000
Remote Data Analysis Evaluator: AI Quality & Feedback
Remote Data Analysis Evaluator: AI Quality & Feedback

Mercor • Greater London

Remote
GBP 34,000 - 55,000