Novel AI Lead Methodologist

Socket.dev

Washington (District of Columbia)

On-site

USD 171,000 - 247,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

Google is seeking a senior researcher in AI testing and evaluation to design novel testing methodologies for emergent AI, partnering with data science and engineering teams to build scalable infrastructure and experiments.

The role focuses on advancing evaluation frameworks, mitigating model risks, and informing product safety decisions; it involves collaboration across DeepMind and Google’s Data Science teams, with exposure to evolving AI capabilities.

Qualifications

  • Bachelor's degree or equivalent practical experience.
  • 5+ years of data analysis for AI testing or related field.

Responsibilities

  • Drive the methodological frontier of model evaluation with data-driven frameworks for testing.
  • Define testing and safety standards with cross-functional teams and develop mitigations.
  • Lead cross-functional teams to implement safety initiatives and advise leadership on complex safety issues.
  • Represent Google's AI safety efforts in external forums and mentor analysts on adversarial techniques.
  • Work with sensitive content and may encounter graphic or upsetting topics.

Skills

Data analysis
AI testing
Experimentation
Quantitative research
Cross-functional collaboration

Education

Bachelor's degree or equivalent practical experience
Master's degree or PhD (preferred)

Tools

Python
SQL

Job description

Minimum qualifications:
  • Bachelor's degree or equivalent practical experience.
  • 10 years of experience in AI testing or research, data analytics, data science, or a related field.
Preferred qualifications:
  • Master's degree or PhD in relevant field.
  • 5 years of experience in data analysis for AI Testing with experience in SQL or Python.
  • Experience building or partnering with engineering teams to build prototypes for AI testing.
  • Experience in designing and conducting experiments or quantitative research, preferably in a technology or AI context.
  • Experience in AI systems, machine learning, and their potential risks.
  • Strong technical competency with a data-driven investigative approach to solve complex tests, including demonstrable proficiency in data manipulation, analysis, and automation using languages like Python and SQL.
About the job:

Novel Testing is a team within Trust and Safety specializing in complex testing, defining protocols and methodologies for assessing risk where best practices do not currently exist. We pioneer and scale testing programs, streamlining the launch of trustworthy, novel AI products.

Work spans from designing first-of-their-kind evaluations for Google’s most ambitious product bets—including autonomous agents, personalization, and the latest hardware—to developing new methodologies for assessing novel foundational model capabilities as they emerge.

Advancing in AI evaluation is central to this mission. To scale these methods, we partner closely with engineering teams to build the infrastructure and tools required for automated evaluation.

In this role, you will lead the development of novel testing methodologies for emergent AI, designing evaluation frameworks where established standards do not yet exist. You will address complex testing questions with creative experimentation, designing sophisticated prompt strategies and quantitative analyses to identify systemic risks and edge cases in GenAI products.

Bridging the gap between theory and execution, you will move quickly to build and prototype testing solutions that incorporate methodological best practices. You will then partner directly with data science and engineering teams to inform the development of novel testing approaches and automated infrastructure, ensuring your insights scale effectively across Google’s ecosystem. This position demands a researcher’s mindset—capable of deep qualitative and quantitative inquiry—paired with the technical agility to translate those findings into scalable, engineering prototypes.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $171000 - $247000 (USD) + 20% bonus target + equity + benefits

Learn more about benefits at Google.

Responsibilities:
  • Drive the methodological frontier of model evaluation. Partner with DeepMind and Data Science, developing novel, data-driven methodologiesfor structured and unstructured testing of emerging AI products. Move beyond standard benchmarks, designing sophisticated experimental frameworks, uncovering latent model behaviors and capabilities.
  • Define testing and safety standards, working with cross-functional colleagues to ensure they are met. Perform analyses and drive insights to develop model-level and product-level safety mitigations.
  • Lead and influence cross-functional teams to implement safety initiatives. Advise executive leadership on complex safety issues.
  • Represent Google's AI safety efforts in external forums and collaborations, contributing to industry-wide best practices. Mentor analysts, fostering a culture of excellence, acting as a subject matter expert on adversarial techniques.
  • Work with sensitive content or situations and may be exposed to graphic, controversial or upsetting topics or content.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Novel AI Lead Methodologist
Novel AI Lead Methodologist

Google • Austin (TX)

On-site
USD 171,000 - 248,000
Health insurance
Dental insurance
Vision insurance
+4
Novel AI Lead Methodologist
Novel AI Lead Methodologist

Google • Seattle (WA)

On-site
USD 171,000 - 247,000
Health insurance
Dental insurance
Vision insurance
+5
Novel AI Lead Methodologist
Novel AI Lead Methodologist

Google • Atlanta (GA)

Hybrid
USD 171,000 - 247,000
Health insurance
Retirement 401(k)
Paid time off
+4
Novel AI Lead Methodologist
Novel AI Lead Methodologist

Google • Washington

On-site
USD 171,000 - 248,000
Health insurance
401(k) with company match
Paid time off 20 days
+4
Novel AI Lead Methodologist
Novel AI Lead Methodologist

Google Inc. • Kirkland (WA)

On-site
USD 171,000 - 247,000
Health insurance
401(k) with company match
Paid time off
Novel AI Lead Methodologist
Novel AI Lead Methodologist

Google • United States

On-site
USD 171,000 - 248,000
Health insurance
401(k) with company match
Paid time off: 20 days per year
+1
AI Evaluation Lead: Novel Testing Methodologies
AI Evaluation Lead: Novel Testing Methodologies

Google • Washington

On-site
USD 171,000 - 248,000
Health insurance
401(k) with company match
Paid time off 20 days
+4
Senior AI Evaluation Lead — Methodology & Testing
Senior AI Evaluation Lead — Methodology & Testing

Google • Atlanta (GA)

Hybrid
USD 171,000 - 247,000
Health insurance
Retirement 401(k)
Paid time off
+4
AI Evaluation Lead for Novel Testing
AI Evaluation Lead for Novel Testing

Google Inc. • Kirkland (WA)

On-site
USD 171,000 - 247,000
Health insurance
401(k) with company match
Paid time off
Research Engineer, Safety Oversight, DeepMind
Research Engineer, Safety Oversight, DeepMind

Google DeepMind • Mountain View (CA)

On-site
USD 174,000 - 253,000