Research Scientist - Robot Learning (VLA / WAM)

Spaitial Ltd.

Greater London

Hybrid

GBP 90,000 - 130,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

SpAItial is recruiting a senior Research Scientist to train policies that turn world models into robotic actions. You will own VLA and WAM end-to-end, from data to execution, including data pipelines, backbones, action representations, and evaluation.

You will advance embodied AI through domain randomization and calibration, pushing the boundaries of 3D perception, control, and sim-to-real transfer with end-to-end training and RL fine-tuning.

Qualifications

  • PhD in robotics, machine learning, or computer vision with a robot learning focus, from the PhD alone or followed by industry experience.
  • Publications at top venues such as (CoRL, RSS, ICRA, IROS or CVPR, ICCV, ECCV, NeurIPS), open-source work, and/or deployed systems.
  • Deep experience with modern robot policy designs (VLA, WAM, diffusion), trained end to end rather than fine-tuned from a released checkpoint.
  • Strong imitation learning fundamentals, and familiarity with RL fine-tuning of pretrained policies.
  • Fluency with VLM backbones and how to adapt them for control.
  • Expert Python and PyTorch, with multi-node distributed training experience (FSDP or equivalent).

Responsibilities

  • Own the training pipeline for vision-language-action (VLA) and world-action models (WAM) end to end, from data to a policy running on a robot.
  • Contribute to setting the technical direction for embodied research at SpAItial.
  • Close the sim-to-real gap through domain randomization, system identification, and calibration, and build evaluation that predicts real-world transfer.
  • Adapt VLM backbones for control: encoder choice and adapter strategies, co-training.
  • Curate and weight the training mix across heterogeneous robot datasets, spanning differing embodiments, action spaces, and sensor setups.
  • Design action representation and decoding, including tokenization, chunking, diffusion, and flow-matching action experts.
  • Build the world-model components that predict future observations conditioned on action.
  • Run post-training: supervised fine-tuning onto target embodiments, and RL for robustness beyond demonstrations.

Skills

Python
PyTorch
Imitation learning
Reinforcement learning
Robot policy design
Distributed training

Education

PhD in robotics / ML / computer vision with robot learning focus

Tools

FSDP
Git
VLM backbones

Job description

SpAItial is pioneering the next generation of World Models, pushing the boundaries of generative AI, computer vision, and the simulation of reality. We are moving beyond 2D pixels to build models that natively understand the physics and geometry of our world. Our mission is to redefine how industries, from robotics and AR/VR to gaming and cinema, generate and interact with physically-grounded 3D environments.

We're seeking a Research Scientist to train the policies that turn a world model into a robot that acts. You will own vision-language-action (VLA) and world-action models (WAM) end to end, starting, including data, backbone, action representation, training runs, and the evaluation that tells us whether a policy is genuinely competent or merely lucky. A world model that understands geometry and physics still doesn't act on its own; the policy is what closes that gap. This is a senior, hands-on research role for someone who has already trained manipulation policies that worked, and who can say precisely why the ones that didn't failed.

Responsibilities
  • Own the training pipeline for vision-language-action (VLA) and world-action models (WAM) end to end, from data to a policy running on a robot.

  • Contribute to setting the technical direction for embodied research at SpAItial.

  • Close the sim-to-real gap through domain randomization, system identification, and calibration, and build evaluation that predicts real-world transfer.

  • Adapt VLM backbones for control: encoder choice and adapter strategies, co-training.

  • Curate and weight the training mix across heterogeneous robot datasets, spanning differing embodiments, action spaces, and sensor setups.

  • Design action representation and decoding, including tokenization, chunking, diffusion, and flow-matching action experts.

  • Build the world-model components that predict future observations conditioned on action.

  • Run post-training: supervised fine-tuning onto target embodiments, and RL for robustness beyond demonstrations.

Key Qualifications
  • A PhD in robotics, machine learning, or computer vision with a robot learning focus, from the PhD alone or followed by industry experience.

  • Publications at top venues such as (CoRL, RSS, ICRA, IROS or CVPR, ICCV, ECCV, NeurIPS), open-source work, and/or deployed systems.

  • Deep experience with modern robot policy designs (VLA, WAM, diffusion), trained end to end rather than fine-tuned from a released checkpoint.

  • Strong imitation learning fundamentals, and familiarity with RL fine-tuning of pretrained policies.

  • Fluency with VLM backbones and how to adapt them for control.

  • Expert Python and PyTorch, with multi-node distributed training experience (FSDP or equivalent).

At SpAItial, we are committed to creating a diverse and inclusive workplace. We welcome applications from people of all backgrounds, experiences, and perspectives. We are an equal opportunity employer and ensure all candidates are treated fairly throughout the recruitment process.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Robotics AI Research Scientist: Train World Models & Policies
Robotics AI Research Scientist: Train World Models & Policies

Spaitial Ltd. • Greater London

Hybrid
GBP 90,000 - 130,000
Research Engineer - 3D World Models
Research Engineer - 3D World Models

SpAItial AI • Greater London

On-site
GBP 40,000 - 60,000
Research Scientist - World Models
Research Scientist - World Models

SpAItial AI • Greater London

On-site
GBP 80,000 - 100,000
VLA Pre-training Engineer - Deep Learning
VLA Pre-training Engineer - Deep Learning

Humanoid • Greater London

On-site
GBP 90,000 - 120,000
Equity and stock options
30+ days off
Private healthcare
+3
Member of Technical Staff (Robotics)
Member of Technical Staff (Robotics)

Reka AI • Greater London

On-site
GBP 50,000 - 80,000
Research Scientist, Robotics RL, DeepMind
Research Scientist, Robotics RL, DeepMind

DeepMind Technologies Limited • Greater London

On-site
GBP 90,000 - 130,000
Research Scientist, Robotics Pre-Training and Data Quality, DeepMind
Research Scientist, Robotics Pre-Training and Data Quality, DeepMind

Google DeepMind • Greater London

On-site
GBP 120,000 - 170,000
Head of Vision – Language – Action (VLA) Development – Manipulation
Head of Vision – Language – Action (VLA) Development – Manipulation

Shaw Daniels Solutions • Greater London

On-site
GBP 190,000 - 230,000
Competitive compensation
Stock options
Paid vacation
+2
VLA Pre-training Engineer (Deep Learning)
VLA Pre-training Engineer (Deep Learning)

Thehumanoid • Greater London

On-site
GBP 70,000 - 100,000
23 days annual leave
Fully funded private healthcare
8% pension contribution
+3
Machine Learning Research Scientist
Machine Learning Research Scientist

Understanding Recruitment • Greater London

On-site
GBP 65,000 - 95,000
Equity