AI Evaluation Specialist - Remote

YO AI Labs

Cambridge

Remote

GBP 21,000 - 41,000

Full time

42 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

YO AI Labs is seeking experienced AI Evaluation Specialists to assess and improve the quality of AI assistant outputs for an enterprise AI training project. You'll evaluate results using defined rubrics for accuracy, relevance, and safety, provide actionable feedback, and help refine evaluation criteria.

This remote role welcomes contributors across multiple regions and time zones. No formal AI experience is required; strong writing, attention to detail, and the ability to work independently on

Qualifications

  • Experience in grading, quality assurance, editorial review, assessment, annotation, or similar analytical work.
  • Advanced, daily use of AI assistants such as ChatGPT, Claude, or similar tools.
  • Strong ability to analyze complex information and communicate findings clearly in writing.
  • Experience with process improvement, rubric development, or quality assessment.
  • Strong critical-thinking skills with a focus on consistency, accuracy, and fairness.
  • Comfortable working independently on high-volume evaluation tasks.
  • Strong attention to detail and ability to identify subtle quality issues.
  • Collaborative approach to resolving ambiguous cases and improving evaluation criteria.

Responsibilities

  • Evaluate AI-generated outputs against detailed rubrics for accuracy, relevance, quality, and adherence to guidelines.
  • Apply consistent and objective judgment across large volumes of examples.
  • Identify reasoning gaps, tool-use issues, logical errors, and inconsistencies.
  • Provide concise, actionable feedback highlighting strengths and areas for improvement.
  • Participate in discussions around rubric interpretation and evolving quality standards.
  • Contribute to process improvements and evaluation best practices.
  • Maintain accurate documentation of evaluations and recommendations.

Skills

Grading
Quality assurance
Annotation
Assessment
Editorial review
Rubric development
Process improvement
Critical thinking

Tools

ChatGPT
Claude

Job description

AI Evaluation Specialist

Role Type: Contractor
Location: Remote — US, Canada, UK, Ireland, Australia, New Zealand

We are seeking experienced AI Evaluation Specialists to assess and improve the quality of AI assistant outputs for an enterprise AI training project.

You'll evaluate AI-generated responses using defined quality standards, identify issues, and provide clear feedback to help improve AI systems. No prior formal AI experience is required.

Scope of Work
  • Evaluate AI-generated outputs against detailed rubrics for accuracy, relevance, quality, and adherence to guidelines.
  • Apply consistent and objective judgment across large volumes of examples.
  • Identify reasoning gaps, tool-use issues, logical errors, and inconsistencies.
  • Provide concise, actionable feedback highlighting strengths and areas for improvement.
  • Participate in discussions around rubric interpretation and evolving quality standards.
  • Contribute to process improvements and evaluation best practices.
  • Maintain accurate documentation of evaluations and recommendations.
Preferred Qualifications
  • Experience in grading, quality assurance, editorial review, assessment, annotation, or similar analytical work.
  • Advanced, daily use of AI assistants such as ChatGPT, Claude, or similar tools.
  • Strong ability to analyze complex information and communicate findings clearly in writing.
  • Experience with process improvement, rubric development, or quality assessment.
  • Strong critical-thinking skills with a focus on consistency, accuracy, and fairness.
  • Comfortable working independently on high-volume evaluation tasks.
  • Strong attention to detail and ability to identify subtle quality issues.
  • Collaborative approach to resolving ambiguous cases and improving evaluation criteria.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Evaluation Specialist - Remote
AI Evaluation Specialist - Remote

YO AI Labs • Oxford

Remote
GBP 42,000 - 83,000
AI Evaluation Specialist - Remote
AI Evaluation Specialist - Remote

YO AI Labs • Manchester

Remote
GBP 34,000 - 48,000
AI Evaluation Specialist - Remote
AI Evaluation Specialist - Remote

YO AI Labs • Greater London

Remote
GBP 55,000 - 92,000
AI Quality Evaluator - Remote Contractor
AI Quality Evaluator - Remote Contractor

YO AI Labs • Manchester

Remote
GBP 34,000 - 48,000
AI Quality Evaluator — Remote Contractor
AI Quality Evaluator — Remote Contractor

YO AI Labs • Oxford

Remote
GBP 42,000 - 83,000
AI Quality Evaluator — Remote Contractor
AI Quality Evaluator — Remote Contractor

YO AI Labs • Cambridge

Remote
GBP 21,000 - 41,000
AI Agent Power User - Remote
AI Agent Power User - Remote

YO AI Labs • Oxford

Remote
GBP 63,000 - 99,000
AI Agent Power User - Remote
AI Agent Power User - Remote

YO AI Labs • Greater London

Remote
GBP 101,000 - 138,000
AI Agent Power User - Remote
AI Agent Power User - Remote

YO AI Labs • Cambridge

Remote
GBP 63,000 - 125,000
Remote AI Quality Evaluator — Contractor
Remote AI Quality Evaluator — Contractor

YO AI Labs • Greater London

Remote
GBP 55,000 - 92,000