A complete application in a minute — tailored resume and cover letter, ready to send.
AuraOne is seeking a Manipulation Risk Risk Evaluator for a remote, contractor role focused on stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document failures, and pair each jailbreak with the violated rubric clause so the safety team can patch gaps.
Adversarial evaluation helps harden models before they ship to customers. Applicants should push on refusal boundaries, document rigorously, and score defenses across single-turn and multi-turn
Manipulation Risk Risk Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Category: AI Safety & Red Teaming · Pay: $65–$70 / hr · Location: Remote — US-eligible · Contractor
Manipulation Risk Risk Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts.
Manipulation Risk Risk Evaluator is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.
Push on refusal boundaries and dual-use risk before a model ships.
Hourly rate confirmed after the interview process.
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.