Senior AI Red Team Scientist for Safety & Evaluation

SupportFinity™

New York (NY)

On-site

USD 190,000 - 211,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

401(k) plan
Bonus program
Equity opportunity

Job summary

Uber is seeking a Senior Applied Scientist to join the AI Red Teaming efforts, focusing on adversarial evaluation, failure analysis, and risk discovery in AI models and agents.

You will design experiments and evaluation frameworks to surface unsafe behaviors, including prompt injection, jailbreaking, and memory poisoning. This role emphasizes safety, robustness, and rigorous scientific methods in real-world AI systems.

Qualifications

  • 5+ years of experience as a Data Scientist, Applied Scientist, or ML Scientist.
  • Hands-on experience with LLMs or generative AI systems.
  • Direct experience with AI red teaming, model safety, or adversarial evaluation.
  • Direct experience with prompt injection, jailbreaks, and LLM failure modes.
  • Strong background in experimental design, evaluation, and statistical analysis.
  • Experience analyzing complex model behavior and failure cases beyond standard metrics.
  • Proficiency in Python and common DS/ML tooling.

Responsibilities

  • Design and execute AI red-teaming experiments against LLMs and AI agents to identify: prompt injection, jailbreaking, policy bypass, model and tool poisoning, context/memory poisoning, behavioral drift, unsafe autonomy.
  • Develop adversarial datasets, probes, and test harnesses to systematically evaluate model and agent behavior under attack.
  • Define and track AI risk metrics beyond accuracy (e.g., failure rates, drift indicators, unsafe action likelihood, confidence miscalibration).
  • Analyze agent workflows and decision traces to understand how failures emerge across multi-step reasoning and tool use.
  • Collaborate with security engineers and AI platform teams to translate findings into guardrails, mitigations, and design improvements.
  • Build reusable evaluation pipelines to support continuous red teaming and regression testing as models and agents evolve.

Skills

5+ years experience
LLMs / Generative AI
AI red teaming
Prompt injection / jailbreaks
Experimental design
Python

Job description

Uber is seeking a Senior Applied Scientist to join the AI Red Teaming efforts, focusing on adversarial evaluation, failure analysis, and risk discovery in AI models and agents.

You will design experiments and evaluation frameworks to surface unsafe behaviors, including prompt injection, jailbreaking, and memory poisoning. This role emphasizes safety, robustness, and rigorous scientific methods in real-world AI systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Applied Scientist - AI Red Teaming & Model Risk
Senior Applied Scientist - AI Red Teaming & Model Risk

SupportFinity™ • New York (NY)

On-site
USD 190,000 - 211,000
401(k) plan
Bonus program
Equity opportunity
Cyber Red Team Specialist: AI Safety & Adversary Testing
Cyber Red Team Specialist: AI Safety & Adversary Testing

OpenAI • Washington

Hybrid
USD 180,000 - 280,000
Relocation assistance
Hybrid work model
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • New York (NY)

On-site
USD 150,000 - 190,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • New York (NY)

On-site
USD 170,000 - 260,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • San Francisco (CA)

On-site
USD 150,000 - 230,000
Staff AI Safety Engineer - Red Team & Guardrails
Staff AI Safety Engineer - Red Team & Guardrails

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive medical, dental, vision, life, and disability insurance
Fully paid parental leave
+2
Remote AI Safety Red Teamer (Contract)
Remote AI Safety Red Teamer (Contract)

United States Digital Space LLC • United States

Remote
USD 96,000 - 116,000
Remote AI Red Team - Safety & Adversarial Testing
Remote AI Red Team - Safety & Adversarial Testing

Obsidian • San Francisco (CA)

On-site
USD 100,000 - 140,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • New York (NY)

On-site
USD 110,000 - 180,000
AI Safety Red Team Engineer
AI Safety Red Team Engineer

Anthropic • San Francisco (CA)

Hybrid
USD 320,000 - 405,000