AI Evaluation Specialist - Remote

YO AI Labs

Berlin

Remote

EUR 31.000 - 55.000

Teilzeit

vor 16 Stunden
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Erhalte eine Antwort von diesem Arbeitgeber — ein Lebenslauf und ein Anschreiben, die genau auf die Eigenschaften eingehen, die gesucht werden.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

YO AI Labs is seeking experienced AI Evaluation Specialists to assess and improve the quality of AI assistant outputs for enterprise training projects. This contractor role is remote and global. You will evaluate responses against rubrics, identify issues, and provide actionable feedback to enhance AI systems.

No formal AI background required, but strong, careful writing is essential. Responsibilities include documenting evaluations, contributing to rubric refinement, and collaborating to

Qualifikationen

  • Experience in grading, quality assurance, editorial review, assessment or similar analytical work.
  • Advanced, daily use of AI assistants such as ChatGPT, Claude, or similar tools.
  • Strong ability to analyze complex information and communicate findings clearly in writing.
  • Experience with process improvement, rubric development, or quality assessment.
  • Strong critical-thinking skills with a focus on consistency, accuracy, and fairness.
  • Comfortable working independently on high-volume evaluation tasks.
  • Strong attention to detail and ability to identify subtle quality issues.
  • Collaborative approach to resolving ambiguous cases and improving evaluation criteria.

Aufgaben

  • Evaluate AI-generated outputs against detailed rubrics for accuracy, relevance, quality, and guidelines.
  • Apply consistent and objective judgment across large volumes of examples.
  • Identify reasoning gaps, tool-use issues, logical errors, and inconsistencies.
  • Provide concise, actionable feedback highlighting strengths and areas for improvement.
  • Participate in discussions around rubric interpretation and evolving quality standards.
  • Contribute to process improvements and evaluation best practices.
  • Maintain accurate documentation of evaluations and recommendations.

Kenntnisse

Grading experience
Quality assurance
Editorial review
Analytical thinking
Documentation/writing
Process improvement
Rubric development
Attention to detail
Independent work
Collaborative problem-solving

Tools

ChatGPT
Claude

Jobbeschreibung

AI Evaluation Specialist

Role Type: Contractor
Location: Remote — US, Canada, UK, Ireland, Australia, New Zealand

We are seeking experienced AI Evaluation Specialists to assess and improve the quality of AI assistant outputs for an enterprise AI training project.

You'll evaluate AI-generated responses using defined quality standards, identify issues, and provide clear feedback to help improve AI systems. No prior formal AI experience is required.

Scope of Work
  • Evaluate AI-generated outputs against detailed rubrics for accuracy, relevance, quality, and adherence to guidelines.
  • Apply consistent and objective judgment across large volumes of examples.
  • Identify reasoning gaps, tool-use issues, logical errors, and inconsistencies.
  • Provide concise, actionable feedback highlighting strengths and areas for improvement.
  • Participate in discussions around rubric interpretation and evolving quality standards.
  • Contribute to process improvements and evaluation best practices.
  • Maintain accurate documentation of evaluations and recommendations.
Preferred Qualifications
  • Experience in grading, quality assurance, editorial review, assessment, annotation, or similar analytical work.
  • Advanced, daily use of AI assistants such as ChatGPT, Claude, or similar tools.
  • Strong ability to analyze complex information and communicate findings clearly in writing.
  • Experience with process improvement, rubric development, or quality assessment.
  • Strong critical-thinking skills with a focus on consistency, accuracy, and fairness.
  • Comfortable working independently on high-volume evaluation tasks.
  • Strong attention to detail and ability to identify subtle quality issues.
  • Collaborative approach to resolving ambiguous cases and improving evaluation criteria.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

AI Evaluation Specialist - Remote
AI Evaluation Specialist - Remote

YO AI Labs • München

Remote
EUR 71.000 - 98.000
AI Evaluation Specialist - Remote
AI Evaluation Specialist - Remote

YO AI Labs • Frankfurt

Remote
EUR 31.000 - 49.000
AI Agent Power User - Remote
AI Agent Power User - Remote

YO AI Labs • Berlin

Remote
EUR 74.000 - 117.000
AI Agent Power User - Remote
AI Agent Power User - Remote

YO AI Labs • München

Remote
EUR 74.000 - 147.000
AI Agent Power User - Remote
AI Agent Power User - Remote

YO AI Labs • Frankfurt

Remote
EUR 74.000 - 147.000
AI Model Evaluation Trainer
AI Model Evaluation Trainer

OpenTrain AI, Inc. • Deutschland

Remote
EUR 17.000 - 44.000
Product Manager - Remote
Product Manager - Remote

YO AI Labs • Frankfurt

Remote
EUR 83.000 - 117.000
Product Manager - Remote
Product Manager - Remote

YO AI Labs • München

Remote
EUR 83.000 - 124.000
Product Manager - Remote
Product Manager - Remote

YO AI Labs • Berlin

Remote
EUR 60.000 - 90.000
AI Trainer - Remote
AI Trainer - Remote

YO AI Labs • Berlin

Remote
EUR 28.000 - 55.000