AI Behavior Researcher - Human Impacts

Transluce

San Francisco (CA)

On-site

USD 250,000 - 450,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Transluce, a fast-moving nonprofit research lab in San Francisco, seeks an AI Behavior Researcher to lead automated evaluations of frontier AI systems, with a focus on wellbeing and user autonomy. You will design pipelines, implement user simulators, and develop robust evaluation rubrics while collaborating with regulators and domain experts.

We value rigorous experimental design and strong Python skills, and we invite applicants across experience levels.

Qualifications

  • Experience designing and validating automated AI evaluation methods, such as LLM-as-a-judge systems or multi-turn benchmarks.
  • Expertise in quantitative generative AI evaluation and measurement; ability to operationalize social concepts.
  • Strong Python proficiency for analysis and tooling.
  • Meticulous experimental design with transparency.
  • Ability to communicate effectively with researchers and decision makers.

Responsibilities

  • Develop novel automated evaluations of AI’s impacts on users, including mental health and decision making.
  • Write code to implement and run automated evaluations, such as user simulators or LLM-as-a-judge pipelines.
  • Design methods to improve ecological validity of evaluations for specific populations.
  • Write and revise rubrics to evaluate model behaviors related to wellbeing and decision making.
  • Collaborate with scientists and research engineers to productionize best practices in AI behavioral evaluation.

Skills

Quantitative AI evaluation
Python proficiency
Experimental design
Communication skills

Tools

LLM-as-a-judge pipelines

Job description

Salary range: $250,000 - $450,000/year + benefits

Description: Transluce is a fast-moving nonprofit research lab building the public tech stack for AI evaluation and oversight. We are pioneering research into the behaviors of AI chatbots and their effect on user wellbeing, and we’re improving outcomes for millions of sensitive AI interactions with vulnerable users.

About the role: As an AI Behavior Researcher, you will lead projects to design and develop automated evaluations of frontier AI systems that are technically sophisticated, scientifically valid, and concretely impactful. This includes expanding on our existing evaluation pipelines to conduct novel analyses of AI behaviors that affect the autonomy and wellbeing of specific user groups (e.g., children or users located in countries beyond the United States).

As an early member of a highly collaborative team, you will learn and grow quickly, and work directly with frontier labs to improve AI evaluations design, with regulators to improve independent oversight of AI, and with domain experts and affected populations to enhance the realism and relevance of our evaluations.

Core responsibilities:
  • Develop novel, valid automated evaluations of AI’s impacts on users, including their mental health and decision making.
  • Write code to implement and run automated evaluations, such as user simulators or LLM-as-a-judge pipelines.
  • Design methods to improve the ecological validity and realism of automated evaluations for specific populations, such as customizing existing user simulation methods to capture the vocabulary used by children.
  • Write and revise judge rubrics to evaluate model behaviors related to user wellbeing and decision making, systematizing abstract, socially situated concepts into clear measurement criteria.
  • Collaborate with scientists and research engineers to productionize best practices in AI behavioral evaluation.
Minimum qualifications:
  • Expertise on quantitative generative AI evaluation and measurement. Good intuition about how to systematize and operationalize complex social concepts.
  • Relevant experience designing and validating automated AI evaluation methods, such as LLM-as-a-judge systems or multi-turn benchmarks.
  • Proficiency in Python to implement analysis and evaluation tooling.
  • Meticulous, good experimental design, epistemic self-awareness and transparency.
  • Ability to balance between the needs of AI researchers and domain experts, as well as between researchers and senior decision makers.
  • Strong communication skills, low ego, openness to giving and receiving feedback.
Preferred qualifications (not required):
  • Experience running automated evaluations at scale or in a production context.
  • Experience conducting controlled human subjects experiments to validate automated evaluation methods.
  • Experience in customer-facing, consulting, or forward-deployed roles translating ambiguous stakeholder needs into concrete deliverables.
  • Experience or training in human-centered design or HCI research methods, including working with domain experts or impacted communities.
  • Experience or demonstrated interest in studying AI’s psychological or social impacts, such as for crisis support, manipulation or sycophancy, political persuasion, or displacing human relationships.
  • Experience designing multilingual generative AI evaluations.
  • Experience and comfort using AI coding agents at work.

We are hiring at all levels of experience and would encourage those enthusiastic about the role who do not meet all of the qualifications to apply. We are located in San Francisco and excited to work together in-person. We are open to sponsoring international visas.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Behavior Researcher - Agent Alignment
AI Behavior Researcher - Agent Alignment

Transluce • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 450,000
AI Behavior Researcher - Child Safety and Mental Health
AI Behavior Researcher - Child Safety and Mental Health

Transluce • San Francisco (CA)

On-site
USD 250,000 - 450,000
AI Behavior Engineer
AI Behavior Engineer

Transluce • San Francisco (CA)

On-site
USD 310,000 - 500,000
AI Behavior Researcher — Impactful AI Evaluation Lead
AI Behavior Researcher — Impactful AI Evaluation Lead

Transluce • San Francisco (CA)

On-site
USD 250,000 - 450,000
Frontier AI Behavior Research Scientist
Frontier AI Behavior Research Scientist

Transluce • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 450,000
AI Evaluation Engineer
AI Evaluation Engineer

DeepRec.ai • Denver (CO)

On-site
USD 162,000 - 198,000
Research Engineer - Scalable Interpretability
Research Engineer - Scalable Interpretability

Transluce • San Francisco (CA)

On-site
USD 250,000 - 500,000
Research Engineer, Model Evaluations
Research Engineer, Model Evaluations

Anthropic • New York (NY), San Francisco (CA)

On-site
USD 320,000 - 485,000
Generous vacation and parental leave
Flexible working hours
Lovely office space for collaboration
Research Engineer / Scientist, Alignment
Research Engineer / Scientist, Alignment

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
Research Engineer / Research Scientist - Model Behavior at OpenAI San Francisco, CA
Research Engineer / Research Scientist - Model Behavior at OpenAI San Francisco, CA

OpenAI • San Francisco (CA)

On-site
USD 310,000 - 460,000
Equity offers