An application made for this job — a tailored resume and cover letter that speak straight to the posting.
AI Trainer Jobs is seeking a remote Computer-Use Agent Evaluator to review prompts and model outputs against a versioned rubric. You will label edge cases, provide structured feedback, and help retrain the modeling team.
Responsibilities include evaluating outputs, selecting the stronger answer with rationale, labeling issues, and calibrating scores weekly. This is a remote independent contractor role with hourly compensation after interview.
Computer-Use Agent Evaluator is a remote evaluation track for reviewing computer use agent evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
AI data reviewers help turn computer use agent evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Review frontier model outputs. Judge benchmark failures and calibrate other evaluators.
Role details
Track Evaluation & annotation Work model Remote Independent specialist contractor Compensation Hourly rate confirmed after the interview process. Eligible from US
What you should bring
Role signals
Example tasks
Useful experience
Compensation and schedule
Hourly rate confirmed after the interview process.
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.
Skills used in matching