Research Engineer (AI Safety)

Axiōma Search

Greater London

On-site

GBP 90,000 - 180,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Axiōma Search is exploring frontline safety for agentic AI systems in real-world environments. The role focuses on defining what good agent safety looks like in practice and reducing risk at the agent layer. You will design evaluations, build guardrails, and collaborate with research and engineering to improve safety while maintaining usefulness.

Shortlisted candidates will be contacted within 48 hours. The team seeks someone with strong engineering or research background in AI safety and solid

Qualifications

  • Strong background in AI safety, alignment, evaluations, or LLM systems.
  • Proficient in Python with practical tooling experience.
  • Ability to reason about risk in real product behaviour.
  • Experience designing evaluations or analysing failures.
  • Clear judgment and communicating while adapting to an evolving brief.

Responsibilities

  • Design and run safety evaluations for agent behaviour across realistic tasks and environments
  • Identify failure modes in tool use, planning, browsing, and multi-step execution
  • Build mitigations, guardrails, and intervention strategies around risky agent behaviour
  • Work with research and engineering teams to improve safe behaviour without killing usefulness
  • Turn concrete incidents or near-misses into better tests, policies, and system changes
  • Help define internal frameworks for agent safety, oversight, and operational risk
  • Contribute hands-on to the systems used to evaluate and monitor safety over time

Skills

AI safety engineering
Python programming
risk assessment
evaluations / red-teaming
communication skills

Job description

About

This is a well-funded frontier AI company building agentic systems that automate complex, multi-step work.

The product and research stack are getting stronger, and safety is increasingly about how those systems behave in realistic environments rather than in isolated model evaluations.

Model safety matters. But once systems can browse, call tools, and complete multi-step tasks, a different class of risk appears. This role is about understanding and reducing that risk at the agent layer.

The brief is still forming. That is part of the appeal. They want someone who can help define what good agent safety looks like in practice.

What you’ll do
  • Design and run safety evaluations for agent behaviour across realistic tasks and environments
  • Identify failure modes in tool use, planning, browsing, and multi-step execution
  • Build mitigations, guardrails, and intervention strategies around risky agent behaviour
  • Work with research and engineering teams to improve safe behaviour without killing usefulness
  • Turn concrete incidents or near-misses into better tests, policies, and system changes
  • Help define internal frameworks for agent safety, oversight, and operational risk
  • Contribute hands-on to the systems used to evaluate and monitor safety over time
What you’ll need
  • Strong engineering or research background in AI safety, alignment, evaluations, or LLM systems
  • Good Python skills and comfort building practical tools rather than only writing papers
  • Ability to reason clearly about risk in real product behaviour, not just offline benchmarks
  • Experience designing evaluations, red-teaming systems, or analysing model/agent failures
  • Strong judgment, clear communication, and comfort with an evolving brief
  • Interest in the safety problems that appear once agents start acting in the world
Optional Bonus
  • Experience with policy systems, trust and safety, or security-style threat modelling
  • Familiarity with tool-using agents or computer-use systems

Shortlisted candidates will be contacted within 48 hours.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer — AI Safety & Agent Risk Evaluation
Research Engineer — AI Safety & Agent Risk Evaluation

Axiōma Search • Greater London

On-site
GBP 90,000 - 180,000
Founding Engineer, Agent Systems
Founding Engineer, Agent Systems

TechTree • Greater London

On-site
GBP 60,000 - 90,000
Researcher, Robustness & Safety Training
Researcher, Robustness & Safety Training

OpenAI • Greater London

On-site
GBP 218,000 - 328,000
Research Scientist/Engineer (Agentic Systems)
Research Scientist/Engineer (Agentic Systems)

White Circle • Greater London

Hybrid
GBP 112,000 - 187,000
Relocation package
Medical insurance (France)
All hardware and tools provided
+2
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Mercor • Greater London

On-site
GBP 70,000 - 110,000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Obsidian • Greater London

On-site
GBP 70,000 - 110,000
AI Safety Specialist - Fully Remote | Upto $70/hr
AI Safety Specialist - Fully Remote | Upto $70/hr

Obsidian • Greater London

Remote
GBP 90,000 - 130,000
Researcher, Frontier Risk Mitigations
Researcher, Frontier Risk Mitigations

OpenAI • Greater London

On-site
GBP 218,000 - 328,000
AI Safety Practitioner - Expert Evaluator
AI Safety Practitioner - Expert Evaluator

Mercor • Greater London

On-site
GBP 70,000 - 110,000
Research Scientist - AI Safety
Research Scientist - AI Safety

Faculty • Greater London

Hybrid
GBP 90,000 - 130,000
Unlimited Annual Leave Policy
Private healthcare and dental
Enhanced parental leave
+3