Research Scientist - Agentic Reasoning & RL Systems

DeepMind Technologies Limited

Mountain View (CA)

On-site

USD 207,000 - 300,000

Full time

20 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Equity
Bonus target
Benefits package

Job summary

DeepMind Technologies Limited invites applications for a senior researcher/engineer on the PRISM team. You will work across the full research and engineering lifecycle, developing distributed post-training infrastructure and algorithms that enable Gemini models to solve complex multistep problems autonomously.

We pursue breakthroughs and production-ready capabilities for Gemini and Gemma, with a focus on scalable reasoning, RL scaling, and responsible AI.

Qualifications

  • Bachelor's or Master's in a quantitative field with practical experience.
  • 4+ years building and scaling ML models with deep learning frameworks.
  • Experience in RL, SFT/RLHF/RLAIF, or related agentic areas.

Responsibilities

  • Operate across the full research-and-engineering lifecycle of frontier reasoning and agentic systems.
  • Address unsolved problems in agentic reasoning, turning early prototypes into hardened production features for Gemini releases.
  • Architect and optimize distributed post-training pipelines and agent-environment simulation loops across thousands of accelerators.
  • Design rigorous experiments and failure analyses to isolate performance bottlenecks and communicate findings.
  • Drive technical excellence by maintaining high code quality and architectural health across shared reinforcement learning and modeling codebases.

Skills

ML frameworks
Reinforcement Learning
Distributed systems
Python

Education

Bachelor's/Master's in CS/Math/Physics

Tools

JAX
PyTorch
TensorFlow

Job description

DeepMind Technologies Limited invites applications for a senior researcher/engineer on the PRISM team. You will work across the full research and engineering lifecycle, developing distributed post-training infrastructure and algorithms that enable Gemini models to solve complex multistep problems autonomously.

We pursue breakthroughs and production-ready capabilities for Gemini and Gemma, with a focus on scalable reasoning, RL scaling, and responsible AI.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Scientist/Engineer, Frontier Reasoning, DeepMind
Research Scientist/Engineer, Frontier Reasoning, DeepMind

DeepMind Technologies Limited • Mountain View (CA)

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits package
Head of AI Research & Enterprise Intelligence
Head of AI Research & Enterprise Intelligence

DeepMind Technologies Limited • Mountain View (CA)

On-site
USD 262,000 - 364,000
Generative AI Research Scientist
Generative AI Research Scientist

Google DeepMind • Mountain View (CA)

On-site
USD 147,000 - 210,000
Agentic Research Scientist: Generative AI & RL Expert
Agentic Research Scientist: Generative AI & RL Expert

Google Inc. • Mountain View (CA), Northern (KY)

Hybrid
USD 190,000 - 236,000
Equity
Benefits
Agentic AI Research Engineer — ML & Systems
Agentic AI Research Engineer — ML & Systems

Google DeepMind • Mountain View (CA)

On-site
USD 174,000 - 252,000
Research Scientist, Gemini Safety & Behavior — GenAI
Research Scientist, Gemini Safety & Behavior — GenAI

Google LLC • New York (NY), Mountain View (CA)

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Gemini AI Engineer: RL & Multimodal
Gemini AI Engineer: RL & Multimodal

Google DeepMind • New York (NY)

On-site
USD 174,000 - 252,000
Equity
Post-Training Agentic Research Scientist - Generative AI
Post-Training Agentic Research Scientist - Generative AI

Google • New York (NY)

On-site
USD 174,000 - 252,000
Equity
Bonus target
Benefits
Senior AI Modeling Scientist — LLMs & RL Post-Training
Senior AI Modeling Scientist — LLMs & RL Post-Training

AI Chopping Block, Inc. • Mountain View (CA), Northern (KY)

Hybrid
USD 262,000 - 364,000
Equity
Bonus target
Benefits package
Staff AI Research Scientist — LLM & RL
Staff AI Research Scientist — LLM & RL

DeepMind Technologies Limited • New York (NY)

On-site
USD 207,000 - 300,000
Health insurance
401(k) with company match
Paid time off
+4