Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
AuraOne is seeking a remote Harmful Instruction Refusal Risk Evaluator to stress-test AI systems through adversarial prompts and rigorous documentation. You will craft attack scenarios, record failures with exact steps, and map each breach to the violated policy clause to help patch gaps before release.
The role emphasizes strong written reporting, clear rationale for each attempt, and the ability to work asynchronously for at least 10 hours per week from a US-based remote setup.
AuraOne is seeking a remote Harmful Instruction Refusal Risk Evaluator to stress-test AI systems through adversarial prompts and rigorous documentation. You will craft attack scenarios, record failures with exact steps, and map each breach to the violated policy clause to help patch gaps before release.
The role emphasizes strong written reporting, clear rationale for each attempt, and the ability to work asynchronously for at least 10 hours per week from a US-based remote setup.