Research Scientist, Safety Oversight, DeepMind

Google DeepMind

Mountain View (CA)

On-site

USD 207,000 - 300,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Google DeepMind aims to turn production data into intelligence on the safety of deployed AI models. Safety Oversight is a new team using large-scale production traffic and automated evaluation methods to monitor safety and alignment of deployed models.

We are part of the GenAI Safety organization, collaborating with teams on safety training and evaluation for Gemini and GenMedia models to rapidly mitigate safety risks in production.

Qualifications

  • PhD in Computer Science, a related field, or equivalent practical experience.
  • 3 years of experience building and shipping technical products.
  • Experience in the domain area of generative AI and Large Language Models (LLM).

Responsibilities

  • Build classifiers and large-scale data pipelines to detect model misbehavior and misuse end-to-end.
  • Research and develop cross-context monitoring systems to detect coordinated harms, developing novel signal aggregation methods across disparate user sessions to identify large-scale attack vectors.
  • Think critically about novel methods for monitoring using model activations, actions, chains-of-thought and final answers.
  • Collaborate closely with infrastructure teams and data scientists to scale your work and regularly share results with the wider safety team.

Skills

PhD in CS or related field
Generative AI/LLM experience
3+ years shipping technical products
Collaborative data analysis

Education

PhD in Computer Science or related field
Master’s or PhD in Engineering/CS

Tools

Data pipelines

Job description

  • PhD in Computer Science, a related field, or equivalent practical experience.
  • 3 years of experience building and shipping technical products.
  • Experience in the domain area of generative AI and Large Language Models (LLM).
Minimum qualifications
  • PhD in Computer Science, a related field, or equivalent practical experience.
  • 3 years of experience building and shipping technical products.
  • Experience in the domain area of generative AI and Large Language Models (LLM).
Preferred qualifications
  • Master’s degree or PhD in Engineering, Computer Science, or a related technical field.
  • 3 years of experience developing code, running experiments and analyses collaboratively with coding agents.
  • Experience building large-scale, highly parallelised data pipelines, working on data quality, automated evaluation design and simple statistical modeling.
  • Ability to approach new research questions and implement technical solutions for them at scale.
  • Ability to use AI every day to build and find ways to push the frontier of model capabilities to accelerate work.
About The Job

We aim to turn production data into intelligence on the safety of deployed AI models. Safety Oversight is a new team tasked with using large-scale production traffic and a variety of automated evaluation methods to monitor the safety and alignment of deployed models. Our work will ensure we measure the real-world efficacy of our safety stack—both of in-model safety training and out-of-model safety mitigations to ensure we are effective in our goal of deploying safe models that are used for widespread public benefit.

The Safety Oversight team sits within the GenAI safety organization and is accountable for ensuring that when a model safety issue occurs in production, or when a user is misusing our model at scale, we detect and understand it, so the safety risk can be rapidly mitigated. We will collaborate closely with teams working on safety training and evaluation for Gemini and GenMedia models.

The Generative AI (GenA)I Safety team operates in a fast-paced, highly collaborative environment. We take the possibility of tangibly dangerous model capabilities seriously as AI advances, and we believe that proactive monitoring and deployment-time oversight are critical for safe AI development.

Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $207000 - $300000 (USD) + 20% bonus target + equity + benefits

Responsibilities

Learn more about benefits at Google .

  • Build classifiers and large-scale data pipelines to detect model misbehavior and misuse end-to-end.
  • Research and develop cross-context monitoring systems to detect coordinated harms, developing novel signal aggregation methods across disparate user sessions to identify large-scale attack vectors.
  • Think critically about novel methods for monitoring using model activations, actions, chains-of-thought and final answers.
  • Collaborate closely with infrastructure teams and data scientists to scale your work and regularly share results with the wider safety team.

Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form .

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Scientist, Safety Oversight
Research Scientist, Safety Oversight

Google Inc. • Mountain View (CA)

Hybrid
USD 207,000 - 300,000
Bonus target
Equity
Benefits
Senior Engineering Analyst, AI Answers, Google Search
Senior Engineering Analyst, AI Answers, Google Search

Google • Washington

On-site
USD 159,000 - 230,000
Health insurance
401(k) with company match
Paid time off - 20 days/year
+4
Senior Engineering Analyst, AI Answers, Google Search
Senior Engineering Analyst, AI Answers, Google Search

Google • Kirkland (WA)

Hybrid
USD 159,000 - 230,000
Health benefits
401(k) with company match
PTO 20 days
+4
Senior Engineering Analyst, AI Answers, Google Search
Senior Engineering Analyst, AI Answers, Google Search

Google • Atlanta (GA)

On-site
USD 159,000 - 230,000
Health insurance
Retirement plan
Paid time off
+4
Research Engineer, Responsible Frontier AI Research, DeepMind
Research Engineer, Responsible Frontier AI Research, DeepMind

Google DeepMind • New York (NY)

On-site
USD 207,000 - 300,000
Research Scientist, Gemini Safety and Behavior, DeepMind
Research Scientist, Gemini Safety and Behavior, DeepMind

Google Inc. • New York (NY)

On-site
USD 207,000 - 300,000
Equity
Bonus programme
Benefits
Research Scientist, SAMBA, DeepMind
Research Scientist, SAMBA, DeepMind

Google DeepMind • New York (NY)

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind

Google DeepMind • New York (NY)

Hybrid
USD 207,000 - 300,000
Research Engineer, Responsible Frontier AI Research, DeepMind
Research Engineer, Responsible Frontier AI Research, DeepMind

Google Inc. • New York (NY)

On-site
USD 207,000 - 300,000
Senior Product Manager, AI Model Safety, Google Cloud AI
Senior Product Manager, AI Model Safety, Google Cloud AI

Google Inc. • Sunnyvale (CA)

On-site
USD 192,000 - 278,000