Open-Model AI Safety Lead: Red Team & Validation | Equity

Reflection AI

Greater London

On-site

GBP 100,000 - 180,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Top-tier compensation
Comprehensive medical, dental, and vision insurance
Paid parental leave for all new parents
Paid time off
Daily lunch and dinner provided

Job summary

A cutting-edge AI company in Greater London is seeking an individual with a graduate degree in Computer Science or related fields to lead the red-teaming and adversarial evaluation of their models. The role requires a deep understanding of LLM safety, strong software engineering skills, and experience with Reinforcement Learning. Ideal candidates thrive in a dynamic environment and are passionate about AI advancements. This position offers competitive compensation and comprehensive health benefits.

Qualifications

  • Graduate degree (MS or PhD) in Computer Science, Machine Learning, or equivalent practical experience in AI Safety.
  • Strong software engineering capabilities with experience building automated evaluation pipelines or large-scale ML systems.
  • Experience with Reinforcement Learning (RLHF/RLAIF) and how it impacts model safety.

Responsibilities

  • Own the red-teaming and adversarial evaluation pipeline for models.
  • Work with the Alignment team to ensure models behave reliably under stress.
  • Develop scalable, automated safety benchmarks for model evaluation.

Skills

Deep technical understanding of LLM safety
Software engineering capabilities
Experience with Reinforcement Learning
Adversarial attacks knowledge

Education

Graduate degree in Computer Science or related discipline

Job description

A cutting-edge AI company in Greater London is seeking an individual with a graduate degree in Computer Science or related fields to lead the red-teaming and adversarial evaluation of their models. The role requires a deep understanding of LLM safety, strong software engineering skills, and experience with Reinforcement Learning. Ideal candidates thrive in a dynamic environment and are passionate about AI advancements. This position offers competitive compensation and comprehensive health benefits.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Generative AI Red Team Lead
Generative AI Red Team Lead

ActiveFence • England

On-site
GBP 60,000 - 80,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • Greater London

On-site
GBP 70,000 - 110,000
Frontier AI Safety Red Team Specialist
Frontier AI Safety Red Team Specialist

Mercor • Greater London

On-site
GBP 90,000 - 130,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • Greater London

On-site
GBP 120,000 - 180,000
Remote AI Red Team Specialist - Safety Adversarial Testing
Remote AI Red Team Specialist - Safety Adversarial Testing

Mercor • Greater London

On-site
GBP 60,000 - 90,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Obsidian • Greater London

Remote
GBP 60,000 - 120,000
Frontier AI Safety Red Teamer
Frontier AI Safety Red Teamer

Obsidian • Greater London

Remote
GBP 85,000 - 120,000
AI Safety Red Team Specialist (Remote)
AI Safety Red Team Specialist (Remote)

Mercor • Greater London

On-site
GBP 60,000 - 90,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • Greater London

On-site
GBP 90,000 - 130,000
AI Safety Red Team Expert – Frontier Model Testing
AI Safety Red Team Expert – Frontier Model Testing

Obsidian • Greater London

On-site
GBP 120,000 - 180,000