Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
AuraOne seeks a contractor to join a remote evaluation track for reviewing writer evaluation prompts and responses against its quality rubric. You will compare outputs, label edge cases, and write structured feedback for model retraining.
Responsibilities include evaluating model outputs, tagging issues, and calibrating reviewer quality weekly. Experience with linguistic or trust & safety review is valued, with flexible, async hours from 10 hours per week.
Writer is a remote evaluation track for reviewing writer evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
Category: Frontier Model Evaluation · Pay: Hourly rate confirmed after the interview process · Location: Remote — US-eligible · Contractor
Writer is a remote evaluation track for reviewing writer evaluation prompts and responses against AuraOne's quality rubric.
Writer is a remote evaluation track for reviewing writer evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling team can use to retrain.
AI data reviewers help turn writer evaluation outputs into auditable labels, rationales, and regression cases for AuraOne Human Data.
Review frontier model outputs. Judge benchmark failures and calibrate other evaluators.
Hourly rate confirmed after the interview process.
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.