A complete application in a minute — tailored resume and cover letter, ready to send.
Mercor is seeking an AI Safety Red Teamer for a contract engagement. You design adversarial prompts to stress-test frontier AI models, identify jailbreaks, unsafe behaviors, and policy failures, and document vulnerabilities for safety benchmarking.
In collaboration with AI researchers, you will help improve model alignment, robustness, and safety through rigorous evaluation across sensitive domains such as misinformation, cyber, biosecurity, and political content.
Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .
Position: AI Safety Red Teamer
Type: Contract
Compensation: $70–$84/hour
Location: Remote
Design adversarial prompts to stress-test frontier AI models . Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures. Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains. Document vulnerabilities and contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers to improve model alignment, robustness, and safety.