Content Adversarial Red Team Analyst - English

Jobgether

India

On-site

INR 1,456,000 - 1,985,000

Part time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Compensation: US$13/hr
Freelance contract
Location: India
Language: English
Flexible schedule
Hands-on AI evaluation

Job summary

Jobgether in India is seeking a Content Adversarial Red Team Analyst (English-based) to evaluate AI systems against content safety policies and compliance requirements. You will design challenging test scenarios to explore how AI platforms respond to complex inputs and edge cases.

The freelance role offers potential to transition into full-time and provides hands-on exposure to AI evaluation and safety testing, with responsibilities including documenting scenarios and findings and applying

Qualifications

  • Educational background or equivalent experience in Trust & Safety, Content Safety, Policy, Linguistics, Communications, Journalism, Research, AI Evaluation, or related field.
  • Relevant experience in content safety, trust and safety, AI evaluation, content moderation, policy enforcement, or quality assurance.
  • Strong written English comprehension and communication skills, with the ability to understand nuanced language and context.
  • Strong understanding of user intent, contextual meaning, and the different ways people may communicate.
  • Ability to think creatively and develop challenging, unusual, or unexpected test scenarios.
  • Strong analytical and critical-thinking skills, with the ability to identify patterns, weaknesses, inconsistencies, and behavioral gaps.

Responsibilities

  • Design and execute authorized adversarial test scenarios to evaluate AI systems against defined content safety policies and compliance requirements.
  • Develop diverse, challenging, and unexpected prompts, inputs, and situations designed to test system behavior.
  • Explore edge cases, unusual inputs, complex contexts, and different user behaviors that may expose gaps in safety controls.
  • Evaluate AI-generated responses and identify potential weaknesses, inconsistencies, policy concerns, or failures in enforcement.
  • Test system behavior across different forms of language, context, intent, and communication styles.
  • Identify recurring patterns and scenarios that may require additional investigation, testing, or system improvement.
  • Clearly document test scenarios, system responses, findings, and supporting evidence.
  • Apply project requirements, testing methodologies, and content safety guidelines consistently.
  • Review complex or ambiguous cases and apply sound judgment when evaluating system behavior.
  • Maintain high standards of accuracy, consistency, documentation quality, and attention to detail across assigned evaluations.

Skills

Adversarial testing
AI evaluation
Content safety
Policy enforcement
Documentation

Education

Educational background in Trust & Safety / Content Safety / Policy / Linguistics / Communications / Journalism / Research / AI Evaluation

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for aContent Adversarial Red Team Analyst - Englishbased inIndia.

As a Content Adversarial Red Team Analyst, you will help evaluate AI systems against content safety policies and compliance requirements.
You will design challenging test scenarios that explore how AI platforms respond to complex, unexpected, or ambiguous inputs.
The role involves examining edge cases, testing different forms of language and intent, and identifying gaps in safety controls or policy enforcement.
You will use creativity and analytical judgment to approach systems from different user perspectives and uncover potential weaknesses.
Your findings will help teams understand where AI behavior can be improved and support the development of safer, more reliable systems.
This freelance opportunity offers the potential to transition into a full-time role and provides hands‑on exposure to AI evaluation and safety testing.

Accountabilities:

  • Design and execute authorized adversarial test scenarios to evaluate AI systems against defined content safety policies and compliance requirements.

  • Develop diverse, challenging, and unexpected prompts, inputs, and situations designed to test system behavior.

  • Explore edge cases, unusual inputs, complex contexts, and different user behaviors that may expose gaps in safety controls.

  • Evaluate AI-generated responses and identify potential weaknesses, inconsistencies, policy concerns, or failures in enforcement.

  • Test system behavior across different forms of language, context, intent, and communication styles.

  • Identify recurring patterns and scenarios that may require additional investigation, testing, or system improvement.

  • Clearly document test scenarios, system responses, findings, and supporting evidence.

  • Apply project requirements, testing methodologies, and content safety guidelines consistently.

  • Review complex or ambiguous cases and apply sound judgment when evaluating system behavior.

  • Maintain high standards of accuracy, consistency, documentation quality, and attention to detail across assigned evaluations.

Requirements:
  • Educational background or equivalent experience in Trust & Safety, Content Safety, Policy, Linguistics, Communications, Journalism, Research, AI Evaluation, or a related field.

  • Relevant experience in content safety, trust and safety, AI evaluation, content moderation, policy enforcement, quality assurance, or a related discipline.

  • Strong understanding of content safety principles, policy enforcement, and common safety risks.

  • Strong written English comprehension and communication skills, with the ability to understand nuanced language and context.

  • Strong understanding of user intent, contextual meaning, and the different ways people may communicate.

  • Ability to think creatively and develop challenging, unusual, or unexpected test scenarios.

  • Strong analytical and critical-thinking skills, with the ability to identify patterns, weaknesses, inconsistencies, and behavioral gaps.

  • Comfort working with complex or ambiguous situations and making informed decisions based on defined requirements.

  • Ability to understand and consistently apply detailed testing guidelines and policies.

  • Strong documentation skills and exceptional attention to detail.

  • Familiarity with AI systems, large language models, adversarial testing, red teaming, or AI safety is preferred.

  • Ability to work independently while maintaining consistent quality across a high volume of evaluation tasks.

Benefits:
  • Compensation:US$13 per hour.

  • Contract type:Freelance, with potential opportunity to transition into a full-time role.

  • Location:India.

  • Language:English.

  • Flexibility:Freelance structure offering flexibility in managing assigned work.

  • Professional exposure:Hands‑on experience evaluating AI systems, content safety controls, and policy compliance.

  • Impact:Contribute to identifying safety gaps and improving the reliability of AI-powered systems.

  • Skill development:Opportunity to strengthen expertise in AI evaluation, adversarial testing, trust and safety, and policy analysis.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Content Adversarial Red Team Analyst - English
Content Adversarial Red Team Analyst - English

Welocalize • Delhi

On-site
INR 1,442,000 - 1,966,000
Content Adversarial Red Team Analyst - English India
Content Adversarial Red Team Analyst - English India

Welo Data • India

On-site
INR 1,712,000 - 1,976,000
Trust & Safety Analyst – AI Evaluation - English
Trust & Safety Analyst – AI Evaluation - English

Welocalize • Delhi

On-site
INR 393,000 - 524,000
AdTech Security Analyst - English
AdTech Security Analyst - English

Welocalize • Delhi

On-site
INR 1,573,000 - 2,097,000
AdTech Security Analyst - English India
AdTech Security Analyst - English India

Welo Data • India

On-site
INR 1,521,000 - 2,166,000
AI/ML Evaluator - English
AI/ML Evaluator - English

Jobgether • India

Remote
INR 1,059,000 - 1,456,000
US$9 per hour
Freelance contract
Location: India
+3
Cybersecurity Threat Analyst - English
Cybersecurity Threat Analyst - English

Jobgether • India

On-site
INR 1,323,000 - 1,853,000
US$10/hour
Freelance
Location India
+1
Business Advisory New Associate
Business Advisory New Associate

Accenture in India • Hyderabad

On-site
INR 400,000 - 500,000
Content Evaluation Specialist (Law/ Human Rights)
Content Evaluation Specialist (Law/ Human Rights)

micro1 • New Delhi

On-site
AI Safety Research & Policy Professional
AI Safety Research & Policy Professional

Data Security Council of India • Delhi

On-site
INR 1,200,000 - 1,800,000