Research Engineer: Human Influence & AI Safety

AI Security Institute

Greater London

Hybrid

GBP 65,000 - 145,000

Full time

14 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Pre-release model access
Ample compute
Publication opportunities
Learning stipend
Conference funding

Job summary

AI Security Institute, based in London, seeks a Research Engineer to advance our Human Influence program through scalable RL post-training, evaluation pipelines, and rigorous experimentation on frontier AI safety. You will work alongside researchers and engineers, deploying robust ML systems and contributing to policy-relevant AI governance discussions.

Join a hyper-impactful team with access to pre-release models, ample compute, and opportunities to publish.

Qualifications

  • Proven experience deploying a benchmark, evaluation, or product to users.
  • Clear understanding of the current AI safety literature, and interest in Human Influence topics.
  • Clear and consistent communication.
  • Clear understanding of fundamental Machine Learning concepts.

Responsibilities

  • Designing and building a Reinforcement Learning environment aimed at mitigating model deception in one-to-one conversations or multi-agent threads.
  • Leveraging interpretability methods to identify why models exhibit concerning behavior and designing mitigations.
  • Building the scalable system architecture underpinning repeatable delivery and analysis of model evaluations and benchmarks.
  • Delivering ambitious, engineering-heavy research projects on Human Influence topics using post-training techniques on large compute clusters.

Skills

Reinforcement Learning
Python
PyTorch
JAX
Keras
Docker
Kubernetes

Tools

Docker
Kubernetes
Ray
FastAPI
SLURM

Job description

AI Security Institute, based in London, seeks a Research Engineer to advance our Human Influence program through scalable RL post-training, evaluation pipelines, and rigorous experimentation on frontier AI safety. You will work alongside researchers and engineers, deploying robust ML systems and contributing to policy-relevant AI governance discussions.

Join a hyper-impactful team with access to pre-release models, ample compute, and opportunities to publish.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Frontier AI Safety Research Engineer
Frontier AI Safety Research Engineer

United States Digital Space LLC • Greater London

Hybrid
GBP 65,000 - 145,000
Hybrid work
25 days annual leave
Annual learning stipend
+2
Hybrid Research Engineer — AI Safety & Red Team
Hybrid Research Engineer — AI Safety & Red Team

AI Security Institute (AISI) • Greater London

Hybrid
GBP 65,000 - 145,000
Hybrid working
Annual learning stipend
Conference funding
+1
Frontier AI Safety Research Scientist
Frontier AI Safety Research Scientist

AI Security Institute • Greater London

On-site
GBP 65,000 - 145,000
Hybrid working
Pension contribution
Conference funding
+1
Frontier AI Safety Research Engineer
Frontier AI Safety Research Engineer

Data Science Jobs UK • Greater London

On-site
GBP 70,000 - 120,000
Frontier AI Security Research Engineer
Frontier AI Security Research Engineer

AI Security Institute • Greater London

Hybrid
GBP 65,000 - 145,000
Hybrid working
Central London office access
AI Safety Research Scientist — Build Safe, Trustworthy AI
AI Safety Research Scientist — Build Safe, Trustworthy AI

Faculty • Greater London

Hybrid
GBP 90,000 - 130,000
Unlimited Annual Leave Policy
Private healthcare and dental
Enhanced parental leave
+3
Research Engineer: AI Safety & Alignment
Research Engineer: AI Safety & Alignment

Anthropic • Greater London

Hybrid
GBP 260,000 - 370,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
+1
Senior AI Safety Research Scientist - Red Team (Hybrid)
Senior AI Safety Research Scientist - Red Team (Hybrid)

Faculty AI • Greater London

Hybrid
GBP 90,000 - 120,000
Unlimited Annual Leave Policy
Private healthcare and dental
Enhanced parental leave
+3
Research Engineer, ML Infrastructure for RL Systems
Research Engineer, ML Infrastructure for RL Systems

Anthropic Limited • Greater London

Hybrid
GBP 110,000 - 170,000
Senior Data Scientist - AI Safety & Evaluation Leader
Senior Data Scientist - AI Safety & Evaluation Leader

AI Chopping Block • Greater London

Hybrid
GBP 90,000 - 130,000
Unlimited Annual Leave
Private healthcare
Enhanced parental leave
+3