Research Scientist, Gemini Safety and Behavior, DeepMind

DeepMind Technologies Limited

Mountain View (CA)

Hybrid

USD 207,000 - 300,000

Full time

30 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

DeepMind Technologies Limited’s Alignment and Reliability division seeks a Research Scientist/Engineer to advance novel heuristic and statistical mechanisms for safe, reliable AI systems. You will own end-to-end experiments from conception to live release within a fast, collaborative team.

We pursue ethical integrity across flagship interfaces and APIs, with a strong emphasis on safety, user-benefit, and scientific discovery, including published work and open collaboration.

Qualifications

  • PhD or equivalent practical experience in CS or related field.
  • 2 years of experience in Large Language Model safety or security.

Responsibilities

  • Drive safety and security mitigations for frontier GenAI models at scale in partnership with product areas.
  • Develop red and blue teaming methods for GenAI models including text-to-text, multimodal, and agentic capabilities.
  • Explore data, reasoning and algorithmic solutions to ensure Gemini Models are safe and helpful for everyone.
  • Improve Gemini's adversarial robustness focusing on high-stakes abuse risks in agentic settings.
  • Develop and execute experimental plans to address known gaps or create new capabilities.

Skills

LLM safety
Agentic workflows
Synthetic data pipelines
Research to product
Applied research leadership
Publications at NeurIPS/ICML/ICLR

Education

PhD in Computer Science or related field

Job description

Note: By applying to this position you will have an opportunity to share your preferred working location from the following: New York, NY, USA; Mountain View, CA, USA; London, UK.

Minimum qualifications
  • PhD degree in Computer Science, a related field, or equivalent practical experience.
  • 2 years of experience in Large Language Model safety or security.
Preferred qualifications
  • Experience in developing and leveraging agentic workflows around safety, behavior, and alignment.
  • Experience with synthetic data generation pipelines, building evaluations and mitigations for non-verifiable tasks using methods such as LLM-as-a-judge, rubric-based rewards, etc.
  • Experience taking research from concept to product.
  • Experience with collaborating or leading an applied research project.
  • Strong experimental taste with good judgment regarding baselines, ablations, and what is worth testing.
  • Track record of publications at NeurIPS, ICLR, ICML.
About The Job

The Alignment and Reliability division investigates and creates auditing frameworks, defensive safeguards, specialized toolsets, and autonomous agents to upgrade GDM's flagship foundation systems. The mandate for the Research Scientist/Engineer centers on engineering novel heuristic and statistical mechanisms to elevate user-facing architectures. The operating rhythm is swift and deeply team-oriented, anchored by a collective ethos of mutual assistance, resilience, and camaraderie intervening during urgent escalations while shaping next-generation intelligence.

In this role, you will be a technical contributor capable of advancing fresh investigative hypotheses from conception to live release with full lifecycle accountability, balancing experimentation with rapid incident mitigation.

Our group specializes in enhancing the ethical integrity and alignment standards of machine intelligence. We pioneer foundational infrastructure integrated across major downstream surfaces, including our flagship consumer interfaces, developer APIs, and core search platforms.

Artificial intelligence will be one of humanity's most transformative inventions. At DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort. Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $207000 - $300000 (USD) + 20% bonus target + equity + benefits

Responsibilities

Learn more about benefits at Google .

  • Drive innovation and understanding of safety and security jailbreaks, owning the problem and delivering mitigations which can be deployed at scale in partnership with product areas.
  • Develop red and blue teaming methods for frontier GenAI models spanning text-to-text, multimodal, and agentic capabilities, delivering actionable insights and solutions.
  • Explore data, reasoning and algorithmic solutions to make sure Gemini Models are safe, maximally helpful, and work for everyone.
  • Improve Gemini's adversarial with a focus on high-stakes abuse risks in agentic settings.
  • Develop and execute experimental plans to address known gaps, or construct entirely new capabilities.

Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form .

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Scientist, Gemini Safety and Behavior, DeepMind
Research Scientist, Gemini Safety and Behavior, DeepMind

DeepMind Technologies Limited • New York (NY)

On-site
USD 207,000 - 300,000
Research Scientist, Gemini Safety and Behavior, DeepMind
Research Scientist, Gemini Safety and Behavior, DeepMind

Google • New York (NY)

On-site
USD 207,000 - 300,000
Equity
Bonus target
Research Scientist, Gemini Agent Safety, DeepMind
Research Scientist, Gemini Agent Safety, DeepMind

Google LLC • Mountain View (CA)

On-site
USD 174,000 - 252,000
Equity
Bonus target
Benefits
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind

Google DeepMind • San Francisco (CA)

On-site
USD 207,000 - 300,000
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind

Google DeepMind • New York (NY)

Hybrid
USD 207,000 - 300,000
Research Scientist, Gemini Safety and Behavior, DeepMind
Research Scientist, Gemini Safety and Behavior, DeepMind

Google LLC • New York (NY), Mountain View (CA)

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits
Research Engineer, Cyber Gemini, DeepMind
Research Engineer, Cyber Gemini, DeepMind

Google DeepMind • New York (NY)

On-site
USD 174,000 - 252,000
Equity
Benefits
Research Engineer, Cyber Gemini, DeepMind
Research Engineer, Cyber Gemini, DeepMind

Google DeepMind • Mountain View (CA)

Hybrid
USD 174,000 - 252,000
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind

Google LLC • San Francisco (CA), Mountain View (CA)

On-site
USD 207,000 - 300,000
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind
Research Engineer, Tech Lead Manager, Frontier and AGI Security, DeepMind

Google LLC • California (MO), Northern (KY)

Hybrid
USD 207,000 - 300,000
Equity
Bonus target
Benefits