Turn this role into an interview — a resume and cover letter built around what this employer wants.
AuraOne is seeking AI Safety Experts — English & Thai to remotely stress-test AI systems against adversarial prompts. Reviewers craft attack scenarios, document failure modes, and pair each jailbreak with the violated rubric clause for patching gaps.
The role supports adversarial evaluation to harden models before release, pushing on refusal boundaries and dual-use risk in a governed remote setting. Responsibilities include designing jailbreak prompts, documenting attacks with reproducible
AI Safety Experts — English & Thai is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Category: AI Safety & Red Teaming · Pay: $24–$35 / hr · Location: Remote — US-eligible · Contractor
AI Safety Experts — English & Thai is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.
Push on refusal boundaries and dual-use risk before a model ships.
$24–$35 / hr
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.