Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
AI Trainer Jobs is seeking a Bilingual Hindi Generalist Expert — AI Safety for a remote red-team role that stress-tests AI systems against adversarial prompts. Reviewers craft attack scenarios, document failure modes, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Adversarial evaluation is how the team hardens AI models before they ship to customers.
Bilingual Hindi Generalist Expert — AI Safety is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Category: AI Safety & Red Teaming · Pay: $18–$22 / hr · Location: Remote — US-eligible · Contractor
Bilingual Hindi Generalist Expert — AI Safety is a remote red-team track for stress-testing AI systems against adversarial prompts.
Bilingual Hindi Generalist Expert — AI Safety is a remote red-team track for stress-testing AI systems against adversarial prompts. Reviewers craft attack scenarios, document the failure mode, and pair each successful jailbreak with the rubric clause it violated so the safety team can patch the gap.
Adversarial evaluation is how AuraOne hardens AI models before they ship to customers. Reviewers think like attackers and write up failures with enough rigor that the modeling team can reproduce, fix, and regress-test them.
Push on refusal boundaries and dual-use risk before a model ships.
$18–$22 / hr
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.