AI Alignment Research Engineer - Red Team

AI Security Institute (AISI)

Greater London

Hybrid

GBP 65,000 - 145,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Hybrid working
Learning & development stipend
Conference funding
Pension with 28.97% employer contrib.

Job summary

AI Security Institute (AISI) in London is seeking Research Engineers and Research Scientists to join the Alignment Red Team at junior to principal levels. You will conduct research on misalignment in frontier AI systems with emphasis on loss‑of‑control risks and evaluations.

The role involves building evaluation tools, collaborating with frontier labs, and contributing to reports and publications that advance alignment testing methods for governments and industry.

Qualifications

  • Completed at least one significant research project in AI safety, security or alignment involving engineering, experiment design and analysis on frontier LLMs.
  • Experience writing clean, documented research code for ML experiments.

Responsibilities

  • Research methods to automatically search for misalignment in frontier models, including loss‑of‑control risks like research sabotage and reward‑seeking.
  • Build and run alignment evaluations for loss‑of‑control risks not captured by current benchmarks.
  • Run pre‑deployment evaluations to test AI alignment and report results to frontier AI companies and UK/allied governments.
  • Contribute to public research publications and technical reports advancing misalignment evaluation methodologies.
  • Design and build software/tools, including open‑source, for better alignment evaluations.

Skills

Autonomous research
Engineering experience
Python for ML
Team collaboration
High velocity

Tools

PyTorch

Job description

AI Security Institute (AISI) in London is seeking Research Engineers and Research Scientists to join the Alignment Red Team at junior to principal levels. You will conduct research on misalignment in frontier AI systems with emphasis on loss‑of‑control risks and evaluations.

The role involves building evaluation tools, collaborating with frontier labs, and contributing to reports and publications that advance alignment testing methods for governments and industry.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Alignment Red Team Research Engineer
AI Alignment Red Team Research Engineer

Aisi • Greater London

Hybrid
GBP 65,000 - 145,000
Hybrid work environment
25 days annual leave
Pension contribution 28.97%
Frontier AI Alignment Engineer – Red Team Research
Frontier AI Alignment Engineer – Red Team Research

AI Security Institute • Greater London

On-site
GBP 90,000 - 130,000
Alignment Red Team - Research Engineer/Research Scientist
Alignment Red Team - Research Engineer/Research Scientist

AI Security Institute • Greater London

On-site
GBP 90,000 - 130,000
Research Engineer: AI Safety & Alignment
Research Engineer: AI Safety & Alignment

Anthropic • Greater London

Hybrid
GBP 260,000 - 370,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
+1
Hybrid Research Engineer — AI Safety & Red Team
Hybrid Research Engineer — AI Safety & Red Team

AI Security Institute (AISI) • Greater London

Hybrid
GBP 65,000 - 145,000
Hybrid working
Annual learning stipend
Conference funding
+1
Frontier AI Security Research Engineer
Frontier AI Security Research Engineer

AI Security Institute • Greater London

Hybrid
GBP 65,000 - 145,000
Hybrid working
Central London office access
Frontier AI Safety Research Engineer
Frontier AI Safety Research Engineer

United States Digital Space LLC • Greater London

Hybrid
GBP 65,000 - 145,000
Hybrid work
25 days annual leave
Annual learning stipend
+2
RSI Safety Researcher: Frontline AI Alignment Prep
RSI Safety Researcher: Frontline AI Alignment Prep

OpenAI • Greater London

On-site
GBP 281,000 - 370,000
Frontier AI Safety Research Engineer
Frontier AI Safety Research Engineer

Data Science Jobs UK • Greater London

On-site
GBP 70,000 - 120,000
Frontier AI Safety Research Scientist
Frontier AI Safety Research Scientist

AI Security Institute • Greater London

On-site
GBP 65,000 - 145,000
Hybrid working
Pension contribution
Conference funding
+1