GenAI Safety & Adversarial Testing Specialist

TikTok

San Jose (CA)

On-site

USD 122,000 - 272,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health insurance
Dental insurance
Vision insurance
401(k) with company match
Parental leave
Disability insurance
Life insurance
Wellbeing benefits
Paid holidays (10 per year)
Paid sick days (10 per year)
Paid Personal Time (17 days)

Job summary

TikTok seeks a seasoned Trust & Safety GenAI adversarial tester to uncover emerging risks in our AI models and policies. You will work across product and policy teams to identify failure modes and provide actionable mitigations, before and after launch.

A strong background in AI safety, jailbreak analysis, and cross-functional collaboration is essential. You will investigate prompts, model behavior, and abuse patterns, communicating findings clearly to engineers and policy leads.

Qualifications

  • Minimum 3 years of experience in Trust & Safety, cybersecurity, risk/adversarial testing, or related fields.
  • Experience with prompt testing, jailbreak analysis, LLM evaluation, or adversarial QA.
  • Familiarity with AI safety risks (jailbreaks, hallucinations, bias, misuse patterns).
  • Strong interest in GenAI safety and how AI systems can be compromised under adversarial conditions.
  • Demonstrated ability to independently investigate ambiguous problems and produce clear conclusions.
  • Ability to manage multiple priorities and collaborate across teams.

Responsibilities

  • Conduct structured adversarial testing on AI models, features, and policies to identify risks.
  • Explore product behavior across contexts and user journeys to find unseen failure modes.
  • Investigate jailbreaks, evasions, prompt-based attacks and other adversarial techniques.
  • Document findings with risk descriptions, reproduction steps, severity, and mitigations.
  • Partner with policy, product, and business teams to validate mitigations and root causes.
  • Support development of testing playbooks, taxonomies, and knowledge bases.
  • Stay updated on emerging adversarial trends and external risk landscape.

Skills

Trust & Safety experience
Cybersecurity
Adversarial testing
Prompt testing
LLM evaluation
AI safety risks
Ambiguity investigation
Cross-functional collaboration

Job description

TikTok seeks a seasoned Trust & Safety GenAI adversarial tester to uncover emerging risks in our AI models and policies. You will work across product and policy teams to identify failure modes and provide actionable mitigations, before and after launch.

A strong background in AI safety, jailbreak analysis, and cross-functional collaboration is essential. You will investigate prompts, model behavior, and abuse patterns, communicating findings clearly to engineers and policy leads.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

GenAI Content Red Team Specialist
GenAI Content Red Team Specialist

TikTok • San Francisco (CA)

On-site
USD 130,000 - 272,000
Medical Insurance
Dental Insurance
Vision Insurance
+8
GenAI Safety & Policy Lead
GenAI Safety & Policy Lead

TikTok USDS Joint Venture • San Jose (CA)

On-site
USD 96,000 - 217,000
GenAI Model Assurance Specialist
GenAI Model Assurance Specialist

TikTok USDS Joint Venture • New York (NY)

On-site
USD 99,000 - 196,000
Senior PM: Generative AI Safety & Trust
Senior PM: Generative AI Safety & Trust

TikTok • San Jose (CA)

On-site
USD 186,000 - 374,000
Senior Product Manager - AI Safety & Governance
Senior Product Manager - AI Safety & Governance

TikTok • San Jose (CA)

On-site
USD 186,000 - 374,000
Senior ML Engineer, Risk & Integrity AI Defenses
Senior ML Engineer, Risk & Integrity AI Defenses

TikTok USDS Joint Venture • Seattle (WA)

On-site
USD 178,000 - 342,000
Medical, dental, and vision insurance
401(k) with company match
Parental leave
+4
ML Engineer: AI Safety & Governance Guardrails
ML Engineer: AI Safety & Governance Guardrails

TikTok USDS Joint Venture • Seattle (WA)

On-site
USD 178,000 - 342,000
Medical, Dental & Vision Insurance
401(k) with company match
Parental leave
+1
Frontier AI Policy Scientist - Trust & Safety
Frontier AI Policy Scientist - Trust & Safety

TikTok • San Francisco (CA)

On-site
USD 108,000 - 209,000
Health insurance
401(k) with match
Paid parental leave
+1
AI Content Red Team Analyst - Trust and Safety
AI Content Red Team Analyst - Trust and Safety

TikTok • San Francisco (CA)

On-site
USD 130,000 - 272,000
Medical Insurance
Dental Insurance
Vision Insurance
+8
Senior AI Security Automation Engineer
Senior AI Security Automation Engineer

TikTok USDS Joint Venture • San Jose (CA)

On-site
USD 147,000 - 270,000