AI Safety & Red Teaming Specialist

Weekday 1

United States

Remote

USD 69,000 - 124,000

Part time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Weekday 1 is seeking an experienced AI Safety & Red Teaming Specialist for a remote contract. You will assess adversarial AI risks, design evaluation methodologies, and improve the robustness of large language models against real-world threats.

This role emphasizes ethical hacking, rigorous testing, and transparent reporting. Collaborate with cross-functional teams to translate security findings into concrete recommendations, develop regression test suites, and document testing approaches for

Qualifications

  • 2+ years of experience in AI Safety, Adversarial ML, or related field.
  • Hands-on experience researching prompts, jailbreaks, adversarial attacks, or tool-use.
  • Strong understanding of modern LLM architectures and safety evaluation methodologies.
  • Experience developing security assessments, regression testing, and adversarial evaluation strategies.
  • Excellent analytical and communication skills; ability to explain complex findings clearly.
  • Ability to collaborate within cross-functional teams.

Responsibilities

  • Design and implement advanced evaluation methodologies for AI system safety, including ethical jailbreak testing, prompt injection detection, LLM red teaming, and tool-use abuse scenarios.
  • Develop cross-domain adversarial testing strategies to uncover complex, multi-turn attack patterns and model vulnerabilities.
  • Build, maintain, and enhance regression test suites to continuously assess jailbreak susceptibility and prompt injection risks.
  • Create comprehensive evaluation frameworks that simulate real-world adversarial threats to improve AI robustness and reliability.
  • Collaborate with technical teams to translate security findings into actionable recommendations for AI safety improvements.
  • Document testing methodologies, findings, and best practices through clear technical reports and presentations for both technical and non-technical stakeholders.

Skills

AI Safety
LLM Red Teaming
Prompt Injection
Adversarial Machine Learning

Education

Master's or PhD in CS/Cybersecurity/ML/AI

Job description

This role is for one of our clients

$50- $90/hourpay

Role Title: AI Safety & Red Teaming Specialist
Role Type: Contractor
Location: Remote

We are looking for an experienced AI Safety & Red Teaming Specialist to help evaluate and strengthen the safety, security, and robustness of next-generation AI systems. In this role, you will leverage your expertise in adversarial AI, LLM security, and AI safety to identify vulnerabilities, design evaluation methodologies, and improve the resilience of large language models against real-world threats.

This opportunity is ideal for professionals passionate about AI security, ethical hacking, and developing robust evaluation frameworks for advanced AI systems.

Requirements

Key Responsibilities
  • Design and implement advanced evaluation methodologies for AI system safety, including ethical jailbreak testing, prompt injection detection, LLM red teaming, and tool-use abuse scenarios.
  • Develop cross-domain adversarial testing strategies to uncover complex, multi-turn attack patterns and model vulnerabilities.
  • Build, maintain, and enhance regression test suites to continuously assess jailbreak susceptibility and prompt injection risks.
  • Create comprehensive evaluation frameworks that simulate real-world adversarial threats to improve AI robustness and reliability.
  • Collaborate with technical teams to translate security findings into actionable recommendations for AI safety improvements.
  • Document testing methodologies, findings, and best practices through clear technical reports and presentations for both technical and non-technical stakeholders.
Required Qualifications
  • 2+ years of experience in AI Safety, Adversarial Machine Learning, LLM Red Teaming, AI Security, or a related field.
  • Hands-on experience researching, testing, or identifying vulnerabilities involving prompt injection, ethical jailbreaks, adversarial attacks, or tool-use exploitation.
  • Strong understanding of modern LLM architectures, prompt engineering, and AI safety evaluation methodologies.
  • Experience developing structured security assessments, regression testing frameworks, and adversarial evaluation strategies.
  • Excellent analytical, documentation, and communication skills with the ability to explain complex technical findings clearly.
  • Ability to collaborate effectively within cross-functional technical teams.
Preferred Qualifications
  • Master's or PhD in Computer Science, Cybersecurity, Machine Learning, Artificial Intelligence, or a related discipline.
  • Contributions to AI security research, open-source AI safety tools, conference presentations, or published research.
  • Experience with AI model evaluation frameworks, prompt engineering techniques, and AI security assessment tools.
  • Background in multidisciplinary AI safety, cybersecurity, or adversarial machine learning projects.
Must-Have Skills
  • AI Safety
  • LLM Red Teaming
  • Prompt Injection
  • Adversarial Machine Learning
Good-to-Have Skills
  • Ethical Jailbreaking
  • AI Security
  • Prompt Engineering
  • AI Evaluation Frameworks
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Safety Specialist - Remote | Upto $84/hr
AI Safety Specialist - Remote | Upto $84/hr

Obsidian • New York (NY)

Remote
USD 140,000 - 210,000
AI Safety Specialist - Remote | Upto $84/hr
AI Safety Specialist - Remote | Upto $84/hr

United States Digital Space LLC • United States

Remote
USD 96,000 - 116,000
AI Safety Experts - English & Finnish
AI Safety Experts - English & Finnish

Weekday 1 • United States

On-site
USD 66,000 - 85,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • New York (NY)

On-site
USD 150,000 - 190,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • New York (NY)

On-site
USD 170,000 - 260,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • San Francisco (CA)

On-site
USD 150,000 - 230,000
AI Red Team Engineer for LLMs
AI Red Team Engineer for LLMs

OpenTrain AI, Inc. • United States

Remote
USD 47,000 - 63,000
Remote work
Worldwide eligibility
LLM Red Team Specialist - Failure Modes & Edge Cases
LLM Red Team Specialist - Failure Modes & Edge Cases

Weekday 1 • United States

Remote
USD 83,000 - 124,000
Fully remote
Weekly payments
Remote LLM Safety & Red Teaming Expert
Remote LLM Safety & Red Teaming Expert

Weekday 1 • United States

Remote
USD 69,000 - 124,000
AI Safety Expert - Red Team
AI Safety Expert - Red Team

Mercor • San Francisco (CA)

Remote
USD 23,000 - 34,000