GenAI Safety Red Team Analyst — Trust & Safety

TikTok

San Jose (CA)

On-site

USD 122,000 - 272,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

TikTok's Trust & Safety GenAI & Emerging Products team seeks an experienced professional to conduct structured adversarial testing of AI models, features, and policies, to uncover vulnerabilities and emerging risks.

You will document findings with clear risk descriptions, reproduction steps, severity assessments, and mitigation recommendations, and collaborate with policy, product, engineering, and data science teams to improve safety before and after launch.

Qualifications

  • Minimum 3 years in Trust & Safety, cybersecurity, risk/adversarial testing, or related fields.
  • Experience with prompt testing, jailbreak analysis, LLM evaluation, or adversarial QA.
  • Familiarity with AI safety risks (jailbreaks, hallucinations, bias, misuse patterns).
  • Strong interest in GenAI safety and adversarial risk mitigation.
  • Ability to investigate ambiguous problems and produce evidence-based conclusions.
  • Ability to manage multiple priorities and collaborate across teams.

Responsibilities

  • Conduct structured adversarial testing on AI models, features, and policies to identify vulnerabilities and emerging risks.
  • Explore product behavior across contexts and user journeys to identify model failure modes not captured by standard evaluations.
  • Investigate jailbreaks, evasions, prompt-based attacks, and other adversarial techniques relevant to content safety.
  • Document findings with risk descriptions, reproduction steps, severity assessments, and mitigation recommendations.
  • Partner with cross-functional stakeholders to ensure mitigation validation and root cause closure.
  • Support development of testing playbooks, taxonomies, and internal knowledge bases.
  • Stay updated on emerging adversarial trends and shifts in the external risk landscape.

Skills

Adversarial testing
GenAI safety
LLM evaluation
Jailbreak analysis
Risk assessment
Cross-functional collaboration
Ambiguous problem solving

Tools

Prompt testing tooling

Job description

TikTok's Trust & Safety GenAI & Emerging Products team seeks an experienced professional to conduct structured adversarial testing of AI models, features, and policies, to uncover vulnerabilities and emerging risks.

You will document findings with clear risk descriptions, reproduction steps, severity assessments, and mitigation recommendations, and collaborate with policy, product, engineering, and data science teams to improve safety before and after launch.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GenAI Content Red Team Specialist
GenAI Content Red Team Specialist

TikTok • San Francisco (CA)

On-site
USD 130,000 - 272,000
Medical Insurance
Dental Insurance
Vision Insurance
+8
GenAI Model Assurance Specialist
GenAI Model Assurance Specialist

TikTok USDS Joint Venture • New York (NY)

On-site
USD 99,000 - 196,000
Senior PM: Generative AI Safety & Trust
Senior PM: Generative AI Safety & Trust

TikTok • San Jose (CA)

On-site
USD 186,000 - 374,000
Senior Product Manager - AI Safety & Governance
Senior Product Manager - AI Safety & Governance

TikTok • San Jose (CA)

On-site
USD 186,000 - 374,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • New York (NY)

On-site
USD 150,000 - 190,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • New York (NY)

On-site
USD 170,000 - 260,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • San Francisco (CA)

On-site
USD 150,000 - 230,000
Senior AI Red Team Scientist for Safety & Evaluation
Senior AI Red Team Scientist for Safety & Evaluation

SupportFinity™ • New York (NY)

On-site
USD 190,000 - 211,000
401(k) plan
Bonus program
Equity opportunity
Cyber Red Team Specialist: AI Safety & Adversary Testing
Cyber Red Team Specialist: AI Safety & Adversary Testing

OpenAI • Washington

Hybrid
USD 180,000 - 280,000
Relocation assistance
Hybrid work model
ML Engineer: AI Safety & Real-Time Governance
ML Engineer: AI Safety & Real-Time Governance

TikTok USDS Joint Venture • San Jose (CA)

On-site
USD 137,000 - 360,000