Research Scientist, Safety Oversight, DeepMind

Google LLC

Greater London

On-site

GBP 110,000 - 140,000

Full time

19 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Google DeepMind is seeking a Research Scientist for Safety Oversight to help turn production data into intelligence about the safety of deployed AI models. The Safety Oversight team uses large‑scale production traffic and automated evaluation techniques to monitor real‑world model safety and alignment.

We collaborate with Gemini and GenMedia teams to rapidly mitigate safety risks, leveraging scalable data pipelines and advanced evaluation methods to push model capabilities safely.

Qualifications

  • PhD or equivalent experience in CS or related field.
  • Experience with generative AI and large language models.
  • Experience building and shipping technical projects.

Responsibilities

  • Build classifiers and data pipelines to detect model misbehavior end‑to‑end.
  • Develop cross-context monitoring to identify large‑scale attack vectors.
  • Evaluate monitoring using model activations, actions, and outputs.
  • Collaborate with infrastructure and data science teams to scale work.

Skills

Generative AI
LLMs
Experimentation

Education

Master’s or PhD in Engineering or CS

Tools

Data pipelines

Job description

Research Scientist, Safety Oversight, DeepMind

Share Research Scientist, Safety Oversight, DeepMind

corporate_fare DeepMind place London, UK

  • PhD in Computer Science, a related field, or equivalent practical experience.
  • Experience in the domain area of generative AI and Large Language Models (LLM).
  • Experience building and shipping technical products.
Preferred qualifications:
  • Master’s degree or PhD in Engineering, Computer Science, or a related technical field.
  • 3 years of experience developing code, running experiments and analyses collaboratively with coding agents.
  • Experience building highly parallelised data pipelines, working on data quality, automated evaluation design and simple statistical modeling.
  • Proven ability in approaching new research questions and implementing technical solutions for them at scale.

Ability to use AI every day to build and find ways to push the frontier of model capabilities to accelerate work.

About the job

We aim to turn production data into intelligence on the safety of deployed AI models. Safety Oversight is a new team tasked with using large-scale production traffic and a variety of automated evaluation methods to monitor the safety and alignment of deployed models. Our work will ensure we measure the real‑world efficacy of our safety stack—both of in‑model safety training and out‑of‑model safety mitigations to ensure we are effective in our goal of deploying safe models that are used for widespread public benefit.

The Safety Oversight team sits within the GenAI safety organization and is accountable for ensuring that when a model safety issue occurs in production, or when a user is misusing our model at scale, we detect and understand it, so the safety risk can be rapidly mitigated. We will collaborate closely with teams working on safety training and evaluation for Gemini and GenMedia models.

The Generative AI (GenA)I Safety team operates in a fast‑paced, highly collaborative environment. We take the possibility of tangibly dangerous model capabilities seriously as AI advances, and we believe that proactive monitoring and deployment‑time oversight are critical for safe AI development.

Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high‑quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

About the job

We aim to turn production data into intelligence on the safety of deployed AI models. Safety Oversight is a new team tasked with using large‑scale production traffic and a variety of automated evaluation methods to monitor the safety and alignment of deployed models. Our work will ensure we measure the real‑world efficacy of our safety stack—both of in‑model safety training and out‑of‑model safety mitigations to ensure we are effective in our goal of deploying safe models that are used for widespread public benefit.

The Safety Oversight team sits within the GenAI safety organization and is accountable for ensuring that when a model safety issue occurs in production, or when a user is misusing our model at scale, we detect and understand it, so the safety risk can be rapidly mitigated. We will collaborate closely with teams working on safety training and evaluation for Gemini and GenMedia models.

The Generative AI (GenA)I Safety team operates in a fast‑paced, highly collaborative environment. We take the possibility of tangibly dangerous model capabilities seriously as AI advances, and we believe that proactive monitoring and deployment‑time oversight are critical for safe AI development.

Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high‑quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.

Responsibilities
  • Build classifiers and data pipelines to detect model misbehavior and misuse end‑to‑end.
  • Research and develop cross‑context monitoring systems to detect coordinated harms, developing novel signal aggregation methods across disparate user sessions to identify large‑scale attack vectors.
  • Think critically about novel methods for monitoring using model activations, actions, chains‑of‑thought and final answers.
  • Collaborate closely with infrastructure teams and data scientists to scale your work and regularly share results with the wider safety team.

Google is proud to be an equal opportunity and affirmative action employer. We are committed to building a workforce that is representative of the users we serve, creating a culture of belonging, and providing an equal employment opportunity regardless of race, creed, color, religion, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition (including breastfeeding), expecting or parents‑to‑be, criminal histories consistent with legal requirements, or any other basis protected by law. See also Google's EEO Policy , Know your rights: workplace discrimination is illegal , Belonging at Google , and How we hire .

Google is a global company and, in order to facilitate efficient collaboration and communication globally, English proficiency is a requirement for all roles unless stated otherwise in the job posting.

Equity is granted exclusively and discretionarily by Alphabet Inc. on the basis of an agreement concluded between you and Alphabet Inc. Alphabet Inc. is your sole contractual partner with respect to equity grants. GSU grants are not guaranteed, are discretionary, are subject to approval by the Alphabet Inc. board of directors or its delegate, the terms of the relevant Alphabet Inc. stock plan, and your grant agreement. They have no impact on statutory payments. Current or past grants do not confer an acquired right.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Engineer, Responsible Frontier AI Research, DeepMind
Research Engineer, Responsible Frontier AI Research, DeepMind

Google Inc. • Greater London

On-site
GBP 90,000 - 130,000
Research Scientist, AGI Safety and Alignment, DeepMind
Research Scientist, AGI Safety and Alignment, DeepMind

Google LLC • Greater London

On-site
GBP 70,000 - 110,000
Equity grants
Manager, Applied AI Engineering, DeepMind
Manager, Applied AI Engineering, DeepMind

Google Inc. • Greater London

On-site
GBP 150,000 - 210,000
Research Scientist, Multisensor Robotics, DeepMind
Research Scientist, Multisensor Robotics, DeepMind

Google Inc. • Greater London

On-site
GBP 100,000 - 150,000
Research Scientist, Robotics RL, DeepMind
Research Scientist, Robotics RL, DeepMind

Google Inc. • Greater London

On-site
GBP 120,000 - 150,000
Research Engineer, Gemini Omni, DeepMind
Research Engineer, Gemini Omni, DeepMind

Google LLC • Greater London

On-site
GBP 110,000 - 140,000
Research Scientist, Robotics Pre-Training and Data Quality, DeepMind
Research Scientist, Robotics Pre-Training and Data Quality, DeepMind

Google Inc. • Greater London

On-site
GBP 110,000 - 150,000
Research Scientist, FSF Risk Modeling and Governance, DeepMind
Research Scientist, FSF Risk Modeling and Governance, DeepMind

Google Inc. • Greater London

On-site
GBP 152,000 - 220,000
Governance Manager, Frontier AI Safety and Policy, DeepMind
Governance Manager, Frontier AI Safety and Policy, DeepMind

Google • Greater London

On-site
GBP 139,000 - 151,000
Research Engineer, Gemini Omni, DeepMind
Research Engineer, Gemini Omni, DeepMind

Google • Greater London

On-site
GBP 120,000 - 180,000