Get more replies from employers
Send a job-specific resume in minutes.
Mercor in Warsaw seeks experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk topics.
Responsibilities include identifying jailbreaks and policy failures, evaluating robustness across misinformation, cyber and biosecurity, and documenting findings for benchmarking.
Mercor in Warsaw seeks experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk topics.
Responsibilities include identifying jailbreaks and policy failures, evaluating robustness across misinformation, cyber and biosecurity, and documenting findings for benchmarking.