Content Adversarial Red Team Analyst - English(US)

Socket.dev

Town of Texas (WI)

On-site

USD 74,000 - 99,000

Part time

7 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Socket.dev is seeking a Content Adversarial Red Team Analyst to test AI platforms against content safety policies and compliance requirements. You will design and execute challenging test scenarios, craft prompts, and analyze responses to identify gaps in safety controls.

The ideal candidate is curious, analytical, and able to consider multiple user perspectives while documenting findings clearly. This freelance role offers a pathway to longer-term opportunities based on performance.

Qualifications

  • Educational background or equivalent experience in Trust & Safety, Content Safety, Policy, Linguistics, Communications, Journalism, Research, AI Evaluation, or related field.
  • Experience in content safety, trust and safety, AI evaluation, content moderation, policy enforcement, quality assurance, or related areas.
  • Strong understanding of content safety policies, policy enforcement, and common safety risks.
  • Strong understanding of written language, context, intent, and different ways users may communicate.

Responsibilities

  • Design and execute authorized adversarial test scenarios to evaluate AI systems against content safety policies and compliance requirements.
  • Develop diverse and challenging prompts, inputs, and scenarios to test system behaviour.
  • Explore edge cases, unusual inputs, and complex situations that may reveal gaps in safety controls or policy enforcement.
  • Evaluate AI responses and identify potential weaknesses, inconsistencies, or policy-related concerns.
  • Test how systems respond to different forms of language, context, intent, and user behaviour.
  • Identify recurring patterns or scenarios that may require further testing or improvement.
  • Document test scenarios, system responses, findings, and supporting evidence clearly.
  • Apply defined testing guidelines and project requirements consistently.
  • Review complex cases and use sound judgment when assessing system behaviour.
  • Maintain high levels of quality, accuracy, and attention to detail while completing assigned tasks.

Skills

Trust & Safety
Content Safety
Policy Enforcement
Linguistics
Research

Education

Bachelor-like qualification or equivalent experience

Tools

NLP tooling

Job description

Job Overview

We are seeking creative and analytical Content Adversarial Red Team Analysts to test AI platforms and models against content safety policies and compliance requirements.

In this role, you will design and execute challenging test scenarios to identify gaps in how AI systems respond to complex, unusual, or unexpected inputs. You will explore edge cases and different approaches to assess whether platforms and models consistently follow defined safety policies and compliance boundaries.

An ideal candidate is curious, analytical, and able to think from different user perspectives. You should have a strong understanding of content safety and policy and be comfortable exploring complex scenarios to identify weaknesses, inconsistencies, or gaps in system behaviour.

Your work will help identify areas of improvement and support the development of safer and more reliable AI systems.

Project Details
  • Contract Type: Freelance, with the potential to convert to a full-time role.
  • Pay Rate: US$63 per hour
  • Location: United States
  • Language: English
Responsibilities
  • Design and execute authorized adversarial test scenarios to evaluate AI systems against content safety policies and compliance requirements.
  • Develop diverse and challenging prompts, inputs, and scenarios to test system behaviour.
  • Explore edge cases, unusual inputs, and complex situations that may reveal gaps in safety controls or policy enforcement.
  • Evaluate AI responses and identify potential weaknesses, inconsistencies, or policy-related concerns.
  • Test how systems respond to different forms of language, context, intent, and user behaviour.
  • Identify recurring patterns or scenarios that may require further testing or improvement.
  • Document test scenarios, system responses, findings, and supporting evidence clearly.
  • Apply defined testing guidelines and project requirements consistently.
  • Review complex cases and use sound judgment when assessing system behaviour.
  • Maintain high levels of quality, accuracy, and attention to detail while completing assigned tasks.
Required Qualifications
  • Educational background or equivalent experience in Trust & Safety, Content Safety, Policy, Linguistics, Communications, Journalism, Research, AI Evaluation, or a related field.
  • Experience in content safety, trust and safety, AI evaluation, content moderation, policy enforcement, quality assurance, or related areas.
  • Strong understanding of content safety policies, policy enforcement, and common safety risks.
  • Strong understanding of written language, context, intent, and different ways users may communicate.
  • Ability to think creatively and develop challenging, unusual, or unexpected test scenarios.
  • Strong analytical and critical-thinking skills.
  • Ability to identify patterns, weaknesses, inconsistencies, and gaps in system behaviour.
  • Comfortable working with complex or ambiguous situations and making informed decisions based on defined requirements.
  • Ability to understand and consistently apply detailed testing guidelines and policies.
  • Strong written English comprehension and communication skills.
  • Strong documentation skills and attention to detail.
  • Familiarity with AI systems, large language models, adversarial testing, red teaming, or AI safety is preferred.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Safety Adversarial Red Team Analyst
AI Safety Adversarial Red Team Analyst

Socket.dev • Town of Texas (WI)

On-site
USD 74,000 - 99,000
AI Red Teamer (LLM Generalist)
AI Red Teamer (LLM Generalist)

Handshake • Seattle (WA)

On-site
USD 83,000 - 138,000
AI Red Teamer (LLM Generalist)
AI Red Teamer (LLM Generalist)

Apply • Seattle (WA), Northern (KY)

Hybrid
USD 120,000 - 180,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • San Francisco (CA)

On-site
USD 150,000 - 230,000
Trust & Safety Analyst – AI Evaluation - English(US)
Trust & Safety Analyst – AI Evaluation - English(US)

Socket.dev • Town of Texas (WI)

On-site
USD 30,000 - 36,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • New York (NY)

On-site
USD 170,000 - 260,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • New York (NY)

On-site
USD 150,000 - 190,000
AI Safety Specialist - Fully Remote | Upto $62/hr
AI Safety Specialist - Fully Remote | Upto $62/hr

Visa Hunt • Town of Belgium (WI)

Remote
USD 66,000 - 85,000
AI Adversarial Specialist - Fully Remote | Upto $22/hr
AI Adversarial Specialist - Fully Remote | Upto $22/hr

Visa Hunt • San Francisco (CA)

Hybrid
CAD 38,000 - 42,000
AI Security Specialist (English & Chinese) | $50.5/hr Remote
AI Security Specialist (English & Chinese) | $50.5/hr Remote

Crossing Hurdles • United States

Remote