Remote Ground Truth QA Evaluator & Rubric Specialist

AI Trainer Jobs

United States

Remote

USD 110,000 - 165,000

Part time

5 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

AuraOne seeks a Ground Truth QA Data Specialist to evaluate ground truth QA data prompts and responses against our quality rubric. You will compare model outputs, label edge cases, and draft structured feedback for the modeling team to retrain.

Responsibilities include evaluating outputs against a versioned rubric, tagging edge cases and unsafe content, calibrating reviewer scores against gold standards, and providing clear rationales to improve future data quality.

Qualifications

  • Prior evaluation, annotation, or human-rater experience on ground truth QA data or adjacent content.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning naming the issue and the rubric clause applied.
  • Strong attention to detail and ability to flag prompt issues.
  • Reliable async availability for at least 10 hours per week.

Responsibilities

  • Evaluate ground truth QA data evaluation model outputs against a versioned rubric and assign severity tags.
  • Compare paired responses and select the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibration against gold-standard examples weekly.

Skills

Evaluation experience
Annotation experience
Inter-rater calibration
Attention to detail
Async availability

Job description

AuraOne seeks a Ground Truth QA Data Specialist to evaluate ground truth QA data prompts and responses against our quality rubric. You will compare model outputs, label edge cases, and draft structured feedback for the modeling team to retrain.

Responsibilities include evaluating outputs against a versioned rubric, tagging edge cases and unsafe content, calibrating reviewer scores against gold standards, and providing clear rationales to improve future data quality.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Ground Truth QA Data Specialist
Ground Truth QA Data Specialist

AI Trainer Jobs • United States

Remote
USD 110,000 - 165,000
Remote Image-Text Grounding QA Specialist
Remote Image-Text Grounding QA Specialist

AI Trainer Jobs • United States

Remote
USD 47,000 - 63,000
Remote work
Contractor role
Hourly pay after interview
Remote Rubric Evaluation Specialist
Remote Rubric Evaluation Specialist

AI Trainer Jobs • United States

Remote
USD 28,000 - 55,000
Senior AI Data Quality Evaluator - Remote
Senior AI Data Quality Evaluator - Remote

AI Trainer Jobs • United States

Remote
USD 34,000 - 55,000
Remote Visual Grounding Evaluator
Remote Visual Grounding Evaluator

AI Trainer Jobs • United States

Remote
USD 25,000 - 39,000
Remote Point Cloud Evaluator - QA & Feedback Specialist
Remote Point Cloud Evaluator - QA & Feedback Specialist

AI Trainer Jobs • United States

Remote
USD 34,000 - 48,000
Remote Human Data QA & Evaluation Reviewer
Remote Human Data QA & Evaluation Reviewer

AI Trainer Jobs • United States

Remote
USD 34,000 - 55,000
Remote work
Contractor role
Flexible hours
Remote Rubrics QA Expert — AI Workflow Evaluator
Remote Rubrics QA Expert — AI Workflow Evaluator

AI Trainer Jobs • United States

Remote
USD 48,000 - 69,000
Remote Rubric Ontology QA Specialist
Remote Rubric Ontology QA Specialist

AI Trainer Jobs • United States

Remote
USD 110,000 - 165,000
Remote Unit Test Workflow Reviewer: QA & Rubric Auditor
Remote Unit Test Workflow Reviewer: QA & Rubric Auditor

AI Trainer Jobs • United States

Remote
USD 34,000 - 55,000