Get more replies from employers
Send a job-specific resume in minutes.
Google is seeking a leader in AI testing within its Novel Testing team under Trust and Safety. You will design and implement novel evaluation methodologies for emergent AI, working with data science and engineering to scale automated evaluation across Google’s ecosystem.
This role emphasizes quantitative research, prototype development, and collaboration with cross-functional teams to mitigate risks in GenAI products. StrongPython/SQL skills and industry experience are required.
In accordance with Washington state law, we are highlighting our comprehensive benefits package, which is available to all eligible US based employees. Benefits for this role include:
Note: By applying to this position you will have an opportunity to share your preferred working location from the following: Washington D.C., DC, USA; Atlanta, GA, USA; Austin, TX, USA; Kirkland, WA, USA; Seattle, WA, USA.
Novel Testing is a team within Trust and Safety specializing in complex testing, defining protocols and methodologies for assessing risk where best practices do not currently exist. We pioneer and scale testing programs, streamlining the launch of trustworthy, novel AI products.
Work spans from designing first-of-their-kind evaluations for Google’s most ambitious product bets—including autonomous agents, personalization, and the latest hardware—to developing new methodologies for assessing novel foundational model capabilities as they emerge.
Advancing in AI evaluation is central to this mission. To scale these methods, we partner closely with engineering teams to build the infrastructure and tools required for automated evaluation.
In this role, you will lead the development of novel testing methodologies for emergent AI, designing evaluation frameworks where established standards do not yet exist. You will address complex testing questions with creative experimentation, designing sophisticated prompt strategies and quantitative analyses to identify systemic risks and edge cases in GenAI products.
Bridging the gap between theory and execution, you will move quickly to build and prototype testing solutions that incorporate methodological best practices. You will then partner directly with data science and engineering teams to inform the development of novel testing approaches and automated infrastructure, ensuring your insights scale effectively across Google’s ecosystem. This position demands a researcher’s mindset—capable of deep qualitative and quantitative inquiry—paired with the technical agility to translate those findings into scalable, engineering prototypes.
Individual pay is determined by factors including job-related skills, experience, and relevant education or training.
US: $171,000 - $248,000 (USD) + 20% bonus target + equity + benefits
Learn more about benefits at Google.
Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form.