AI Evaluation Specialist | $35/hr Remote

Crossing Hurdles

Philippines

On-site

PHP 334,800 - 468,720

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Crossing Hurdles is seeking a meticulous evaluator to design and implement concrete AI evaluation tasks with clear prompts, rubrics, and supporting materials. You will observe AI behavior, craft precise English reports, and collaborate with teams to refine benchmarking methods.

The role emphasizes strong writing, attention to detail, and structured assessment practices. You will contribute to robust evaluation frameworks, document findings succinctly, and help drive continuous improvement in AI

Qualifications

  • Have experience in roles emphasizing written precision and structured thinking (paralegal, executive assistant, junior analyst, librarian, QA analyst).
  • Native or fluent English writing; able to produce concise, specific, unambiguous observations.
  • Proven skill designing or applying rubric-based evaluation and scoring frameworks.

Responsibilities

  • Design and implement self-contained evaluation tasks, including prompts, supporting files, and detailed grading rubrics.
  • Meticulously observe AI agent behaviors and produce precise summaries and reports in high-quality English.
  • Iterate and refine evaluation tasks and rubrics based on feedback and teamwork to ensure robust benchmarking.
  • Collaborate with the customer’s team to share insights and drive continuous improvement in evaluation techniques.

Skills

Written precision
Structured thinking
Rubric design
Attention to detail
Clear communication

Job description

  • Design and implement self-contained evaluation tasks, including prompts, supporting files, and detailed grading rubrics to assess AI performance on practical computer-based workflows.
  • Define clear, unambiguous written criteria for successful and unsuccessful task completion across diverse administrative and workflow scenarios.
  • Meticulously observe and document AI agent behaviors, producing crisp, precise summaries and reports in high-quality English.
  • Iterate and refine evaluation tasks and rubrics based on feedback and team collaboration to ensure robust benchmarking methodologies.
  • Collaborate with the customer's team to share insights and help drive continuous improvement in AI evaluation techniques.
Requirements
  • Have a minimum of of experience in roles emphasizing written precision and structured thinking, such as paralegal, executive assistant, junior analyst, librarian, document archival specialist, research assistant, technical writer, or QA analyst.
  • Be native or fluent in English writing, with a demonstrated ability to produce observations that are succinct, specific, and unambiguous.
  • Proven skill in designing or applying rubric-based evaluation, grading against set criteria, or building structured scoring frameworks is required.
  • Possess high attention to detail and the ability to notice subtle patterns or inconsistencies others might miss.
  • Exceptional written and verbal communication skills are necessary, especially for documenting nuanced observations and feedback.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Evaluator | $30/hr Remote
AI Evaluator | $30/hr Remote

Crossing Hurdles • Philippines

On-site
AI Quality Assurance Specialist | $30/hr Remote
AI Quality Assurance Specialist | $30/hr Remote

Crossing Hurdles • Philippines

On-site
PHP 180,000 - 320,000
Remote AI Evaluation Specialist & Rubric Designer
Remote AI Evaluation Specialist & Rubric Designer

Crossing Hurdles • Philippines

On-site
AI Content Evaluator (Remote)
AI Content Evaluator (Remote)

Innodata Inc. • Philippines

On-site
PHP 240,000 - 360,000
AI Agent Evaluation Engineer — Test Architect
AI Agent Evaluation Engineer — Test Architect

Mindrift • Philippines

On-site
Remote AI Language Quality Specialist
Remote AI Language Quality Specialist

YO IT Consulting • Philippines

On-site
Remote AI Evaluation & Calibration Specialist
Remote AI Evaluation & Calibration Specialist

Crossing Hurdles • Philippines

On-site
Remote AI Training & Evaluation Specialist
Remote AI Training & Evaluation Specialist

Alignerr • Philippines

On-site
PHP 424,000 - 763,000
Fully remote
Flexible hours
Remote AI Search Quality Evaluator
Remote AI Search Quality Evaluator

Alignerr • Philippines

On-site
PHP 509,000 - 1,017,000
AI QA Specialist for Frontier LLMs
AI QA Specialist for Frontier LLMs

Crossing Hurdles • Philippines

On-site
PHP 180,000 - 320,000