Stand out for this role — generate a tailored resume and cover letter in about a minute.
AI Safety Experts — English & Marathi is a remote contractor track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Adversarial evaluation helps harden models before shipping. Reviewers think like attackers, write up failures with rigor, ensure reproducibility, and provide patches for the modeling team to implement and
AI Safety Experts — English & Marathi is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Category: AI Safety & Red Teaming · Pay: $20–$22 / hr · Location: Remote — US-eligible · Contractor
AI Safety Experts — English & Marathi is a remote red-team track for stress-testing AI systems against adversarial prompts.
AI Safety Experts — English & Marathi is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.
Push on refusal boundaries and dual-use risk before a model ships.
$20–$22 / hr
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.