Machine Learning Research Scientist

AIXI Labs

Bromley

On-site

GBP 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

AIXI Labs is a small, London-based non-profit doing AI safety research grounded in algorithmic information theory. You’d join a three-person core team—Cole Wyeth (Founder & Executive Director), Marcus Hutter (Research Director), and Aram Ebtekar (Founding Research Scientist)—at an early stage.

You’d help build our empirical research capability essentially from scratch. Minimum qualifications focus on a strong mathematical background and the ability to design and run experiments with modern

Qualifications

  • Mathematical background to rapidly upskill on algorithmic information theory, AIXI, Bayesian statistics, and reinforcement learning theory.
  • Proven ability to independently design, build, and run experiments with modern LLM-based agents.
  • Ability to identify a worthwhile question, and drive it to a concrete result with minimal oversight.
  • Strong verbal and written communication skills.
  • Willingness to work as one of a very small core team, owning problems end-to-end.

Responsibilities

  • Design and run experiments on LLM-based agents that test theoretically-predicted risk factors for loss of control under realistic conditions.
  • Build and evaluate safety mitigations motivated by our theoretical work, and report honestly on where they hold up and where they don’t.
  • Directly help shape the research agenda — with a team this size, your judgment about which experiments are worth running will materially affect what we work on.
  • Contribute to papers, blog posts, or other written output that communicates results to the broader AI safety and ML research communities.

Skills

Bayesian statistics
Reinforcement learning theory
Algorithmic information theory
Experiment design
LLM-based agents
Independent problem solving

Education

PhD in CS or Physics

Job description

London, UK or Berkeley, US

  • Full-time
  • £120,000 – £160,000 / year

AIXI Labs is a small, London-based non-profit doing AI safety research grounded in algorithmic information theory. Most AI research today is either mathematically clean but narrow, or practical but too opaque for rigorous safety analysis. We target a third category: methods general enough to describe powerful agents and provable enough to support real safety claims — using AIXI, the theoretical model of unbounded artificial superintelligence — then port the strongest of these ideas to modern LLM-based agents.

You’d be joining a three-person core team — Cole Wyeth (Founder & Executive Director), Marcus Hutter (Research Director), and Aram Ebtekar (Founding Research Scientist) — at an early stage. We are not an established engineering org, so you’d help build our empirical research capability essentially from scratch.

Minimum Qualifications
  • Mathematical background sufficient to rapidly upskill on algorithmic information theory, AIXI, Bayesian statistics, and reinforcement learning theory (existing expertise on these topics not required).
  • Proven ability to independently design, build, and run experiments with modern LLM-based agents.
  • Ability to identify a worthwhile question, and drive it to a concrete result with minimal oversight.
  • Strong verbal and written communication skills.
  • Willingness to work as one of a very small core team (currently three people), which means owning problems end-to-end, including infrastructure and "glue work" a larger team would delegate.
Preferred Qualifications
  • PhD in computer science, physics, or a related field.
  • Publications at top ML/AI/theory venues (e.g. NeurIPS, ICML, ICLR, ALT, COLT), though we value the strength of your best work over volume.
  • Experience designing empirical tests for safety-relevant agent behaviors: deception, specification gaming, power-seeking, goal misgeneralization, or similar.
  • Background in algorithmic information theory, computability theory, learning theory, or related theoretical computer science.
  • Experience fine-tuning, evaluating, or red-teaming LLM-based and/or RL agents.
Responsibilities
  • Design and run experiments on LLM-based agents that test theoretically-predicted risk factors for loss of control (e.g. deceptive or power-seeking behavior) under realistic conditions.
  • Build and evaluate safety mitigations motivated by our theoretical work, and report honestly on where they hold up and where they don’t.
  • Directly help shape the research agenda — with a team this size, your judgment about which experiments are worth running will materially affect what we work on.
  • Contribute to papers, blog posts, or other written output that communicates results to the broader AI safety and ML research communities.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Safety Research Scientist — ML Theory & Experiments
AI Safety Research Scientist — ML Theory & Experiments

AIXI Labs • Bromley

On-site
GBP 120,000 - 160,000
Research Scientist/Engineer (Agentic Systems)
Research Scientist/Engineer (Agentic Systems)

White Circle • Greater London

Hybrid
GBP 112,000 - 187,000
Relocation package
Medical insurance (France)
All hardware and tools provided
+2
[Expression of Interest] Research Engineer / Scientist, Alignment - London
[Expression of Interest] Research Engineer / Scientist, Alignment - London

Menlo Ventures • Greater London

On-site
GBP 260,000 - 370,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
+1
Research Scientist/Engineer - General Decision & Control Agent
Research Scientist/Engineer - General Decision & Control Agent

Adecco • City Of London

On-site
GBP 70,000 - 120,000
AI Safety Research Scientist — Build Safe, Trustworthy AI
AI Safety Research Scientist — Build Safe, Trustworthy AI

Faculty • Greater London

Hybrid
GBP 90,000 - 130,000
Unlimited Annual Leave Policy
Private healthcare and dental
Enhanced parental leave
+3
Research Engineer / Scientist, Alignment Science - London
Research Engineer / Scientist, Alignment Science - London

Anthropic • Greater London

On-site
GBP 250,000 - 270,000
Competitive compensation
Flexible working hours
Generous vacation and parental leave
Senior/ Principal Research Scientist, AI Safety, Biological/ Physical Sciences
Senior/ Principal Research Scientist, AI Safety, Biological/ Physical Sciences

Lila Sciences • Greater London

Hybrid
GBP 100,000 - 170,000
Autonomous Systems Research Scientist (AI Safety)
Autonomous Systems Research Scientist (AI Safety)

White Circle • Greater London

Hybrid
GBP 112,000 - 187,000
Relocation package
Medical insurance (France)
All hardware and tools provided
+2
ML Research Engineer: Sovereign AI Agent Design & Safety
ML Research Engineer: Sovereign AI Agent Design & Safety

Scale AI • United Kingdom

On-site
GBP 90,000 - 150,000
ML Research Scientist - Member of Technical Staff
ML Research Scientist - Member of Technical Staff

United States Digital Space LLC • Greater London

On-site
GBP 101,000 - 192,000
Competitive salary
Equity ownership
Private healthcare
+2