AI Evaluation Expert - Remote

YO AI Labs

Bengaluru

Remote

INR 3,321,000 - 7,971,000

Part time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

YO AI Labs is seeking experienced AI Evaluation Specialists to assess and improve the quality of AI assistant outputs for an enterprise training project. This contract role focuses on evaluating responses against rubrics, identifying gaps, and providing actionable feedback to improve AI systems.

Prior AI experience is not required, making it accessible to strong analytical writers. You will work remotely on high-volume evaluation tasks, maintain thorough documentation, and engage in rubric

Qualifications

  • Experience in grading, QA, editorial review or similar analytical work.
  • Proficient with AI assistants such as ChatGPT or Claude.
  • Ability to analyze complex information and communicate findings clearly.
  • Experience with rubric development or quality assessment.
  • Attention to detail and consistency under high-volume evaluations.

Responsibilities

  • Evaluate AI-generated outputs against rubrics for accuracy, relevance, quality and guidelines.
  • Apply consistent, objective judgment across large volumes of examples.
  • Identify reasoning gaps, tool-use issues, and inconsistencies.
  • Provide concise, actionable feedback highlighting strengths and improvement areas.
  • Document evaluations and recommendations clearly for stakeholders.
  • Participate in rubric interpretation discussions and evolving quality standards.
  • Contribute to process improvements and best practices in evaluation.

Skills

Grading / QA
Editorial review
Quality assessment
Rubric development
Critical thinking
Independent work
Attention to detail

Tools

ChatGPT
Claude

Job description

AI Evaluation Specialist

Role Type: Contractor
Location: Remote — US, Canada, UK, Ireland, Australia, New Zealand

We are seeking experienced AI Evaluation Specialists to assess and improve the quality of AI assistant outputs for an enterprise AI training project.

You'll evaluate AI-generated responses using defined quality standards, identify issues, and provide clear feedback to help improve AI systems. No prior formal AI experience is required.

Scope of Work
  • Evaluate AI-generated outputs against detailed rubrics for accuracy, relevance, quality, and adherence to guidelines.
  • Apply consistent and objective judgment across large volumes of examples.
  • Identify reasoning gaps, tool-use issues, logical errors, and inconsistencies.
  • Provide concise, actionable feedback highlighting strengths and areas for improvement.
  • Participate in discussions around rubric interpretation and evolving quality standards.
  • Contribute to process improvements and evaluation best practices.
  • Maintain accurate documentation of evaluations and recommendations.
Preferred Qualifications
  • Experience in grading, quality assurance, editorial review, assessment, annotation, or similar analytical work.
  • Advanced, daily use of AI assistants such as ChatGPT, Claude, or similar tools.
  • Strong ability to analyze complex information and communicate findings clearly in writing.
  • Experience with process improvement, rubric development, or quality assessment.
  • Strong critical-thinking skills with a focus on consistency, accuracy, and fairness.
  • Comfortable working independently on high-volume evaluation tasks.
  • Strong attention to detail and ability to identify subtle quality issues.
  • Collaborative approach to resolving ambiguous cases and improving evaluation criteria.
Get your free, confidential resume review.

or drag and drop your file here.