Remote C++ AI Evaluation Auditor (Contractor)

AuraOne

United States

On-site

USD 55,000 - 83,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

AuraOne is seeking remote evaluators for the C++ Systems Programming AI evaluation track. You will review model outputs, apply a versioned rubric, assign severity tags, and provide structured feedback to retrain prompts.

Work is contract-based, US-eligible, with an hourly rate to be confirmed after interview. Strong attention to detail and ability to justify rubric choices are essential for this role.

Qualifications

  • Prior evaluation, annotation, or human-rater experience on C++ systems programming AI evaluation or adjacent content.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the issue and the rubric clause being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.

Responsibilities

  • Evaluate c++ systems programming ai evaluation model outputs against a versioned rubric and assign severity tags for C++ Systems Programming AI Evaluator assignments.
  • Compare paired responses and pick the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them back to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibrating against gold-standard examples each week.
  • Document recurring failure modes so the modeling team can target them in the next training run.

Skills

Model output evaluation
Rubric-based annotation
Severity tagging
Inter-rater calibration
C++ Systems Programming AI evaluation
Software engineering
AI evaluation
Rubric writing
Expert review

Job description

AuraOne is seeking remote evaluators for the C++ Systems Programming AI evaluation track. You will review model outputs, apply a versioned rubric, assign severity tags, and provide structured feedback to retrain prompts.

Work is contract-based, US-eligible, with an hourly rate to be confirmed after interview. Strong attention to detail and ability to justify rubric choices are essential for this role.

Get your free, confidential resume review.
or drag and drop your file here.