Research Engineer, Responsible Frontier AI Research, DeepMind

DeepMind Technologies Limited

Greater London

On-site

GBP 90,000 - 140,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

DeepMind Technologies Limited is seeking a highly adaptable Research Engineer to join the Responsibility team. You will design and scale evaluation frameworks that assess the propensity of frontier language models to engage in harmful manipulation, develop evaluation and mitigation techniques for manipulative model behaviors, and build infrastructure to track model performance across releases.

You will work closely with Research Scientists, safety policy teams, and product stakeholders to ensure

Qualifications

  • Bachelor's degree in Computer Science, Machine Learning, Mathematics, or a related technical field, or equivalent practical experience.
  • 3 years of experience in Python programming.
  • 3 years of experience with ML frameworks such as JAX, PyTorch, or TensorFlow.

Responsibilities

  • Rapidly prototype and deliver scalable engineering solutions across the Responsibility research portfolio.
  • Architect and optimize training and inference pipelines to detect and evaluate harmful manipulation behaviors in frontier language models.
  • Develop post-training strategies to mitigate manipulation risks including deceptive persuasion, sycophancy, and covert influence tactics.
  • Collaborate with research scientists to translate safety research into robust implementations and present results to cross-functional stakeholders.
  • Build and maintain evaluation infrastructure to systematically track model safety performance across releases.

Skills

Python
JAX
PyTorch
TensorFlow

Education

BSc in CS/ML/Math or related field
MS or PhD in CS/Engineering/ML

Tools

C++

Job description

Minimum qualifications:
  • Bachelor's degree in Computer Science, Machine Learning, Mathematics, or a related technical field, or equivalent practical experience.
  • 3 years of experience in Python programming.
  • 3 years of experience with ML frameworks such as JAX, PyTorch, or TensorFlow.
  • Bachelor's degree in Computer Science, Machine Learning, Mathematics, or a related technical field, or equivalent practical experience.
  • 3 years of experience in Python programming.
  • 3 years of experience with ML frameworks such as JAX, PyTorch, or TensorFlow.
Preferred qualifications:
  • Master's degree or PhD in Computer Science, Engineering, or a related field with a focus on machine learning.
  • Experience in Python and C++ for high-performance ML library development.
  • Experience with harmful manipulation detection, persuasion modeling, deceptive behavior analysis, or AI safety evaluation and mitigation.
  • Experience working directly on AI safety, or responsible AI research.
  • Experience building evaluation frameworks, benchmarks, or automated testing pipelines for ML models.
About the job

We are looking for a highly adaptable Research Engineer to join our Responsibility team. In this role, you will design and scale evaluation frameworks that assess the propensity of frontier language models to engage in harmful manipulation, develop evaluation and mitigation techniques for manipulative model behaviors, and build infrastructure to systematically track model performance across releases. You will work closely with Research Scientists, safety policy teams, and product stakeholders to ensure that evaluation results translate into concrete safety improvements.

You will be joining a specialized Applied Research and Responsibility team within Google while closely partnering with teams across Deepmind, Research, Product, and Policy to identify and address key challenges in responsible AI. Our work remains deeply integrated with DeepMind.

Responsibilities
  • Be able to rapidly prototype and deliver scalable engineering solutions across the Responsibility research portfolio.
  • Architect and optimize training and inference pipelines to detect and evaluate harmful manipulation behaviors in frontier language models.
  • Develop post-training strategies to mitigate manipulation risks including deceptive persuasion, sycophancy, and covert influence tactics.
  • Collaborate with research scientists to translate safety research into robust implementations and present results to cross-functional stakeholders.
  • Build and maintain evaluation infrastructure to systematically track model safety performance across releases.

Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form .

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Engineer, Responsible Frontier AI Research
Research Engineer, Responsible Frontier AI Research

Google LLC • Greater London

Hybrid
GBP 90,000 - 150,000
Frontier AI Safety Research Engineer
Frontier AI Safety Research Engineer

Google LLC • Greater London

Hybrid
GBP 110,000 - 170,000
Research Scientist, Safety Oversight, DeepMind
Research Scientist, Safety Oversight, DeepMind

DeepMind Technologies Limited • Greater London

On-site
GBP 120,000 - 180,000
Research Engineer, Responsible AI & Safety Frameworks
Research Engineer, Responsible AI & Safety Frameworks

DeepMind Technologies Limited • Greater London

On-site
GBP 90,000 - 140,000
Research Scientist, Safety Oversight, DeepMind
Research Scientist, Safety Oversight, DeepMind

Google LLC • Greater London

On-site
GBP 110,000 - 140,000
AI Safety and Alignment Researcher
AI Safety and Alignment Researcher

AI Breaking Wire • Greater London

On-site
GBP 90,000 - 150,000
Equity / stock options
Health, wellness & family care
World-class computing resources
+1
Research Scientist, Gemini Safety and Behavior, DeepMind
Research Scientist, Gemini Safety and Behavior, DeepMind

DeepMind Technologies Limited • Greater London

On-site
GBP 120,000 - 180,000
Equity
Benefits
Senior AI Safety Researcher
Senior AI Safety Researcher

AI Breaking Wire • Greater London

On-site
GBP 110,000 - 170,000
Stock options
Health programs and wellness benefits
Learning and development allowances
+1
Research Scientist, Multisensor Robotics, DeepMind
Research Scientist, Multisensor Robotics, DeepMind

Google DeepMind • Greater London

On-site
GBP 120,000 - 180,000
Manager, Applied AI Engineering, DeepMind
Manager, Applied AI Engineering, DeepMind

Google DeepMind • Greater London

On-site
GBP 140,000 - 230,000