AI Behavior Engineer

Transluce

San Francisco (CA)

On-site

USD 310,000 - 500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Transluce, located in San Francisco, seeks an engineer to develop and measure AI model behaviors. In this role, you'll engage with policymakers and civil partners to address key questions regarding AI systems. Responsibilities include building evaluation methods and adapting them for various contexts. The ideal candidate possesses hands-on experience with AI evaluations, strong engineering instincts, and the ability to translate complex needs into actionable solutions. This position is open to sponsoring international visas.

Qualifications

  • Hands-on experience designing and running AI evaluations, particularly behavioral or interactive evaluations.
  • Strong engineering instincts and good judgment about when ‘good enough to ship’ is actually good enough.
  • Experience translating ambiguous stakeholder needs into concrete deliverables.
  • Experience running evaluations at scale or in a production context.
  • Ability to balance needs between researchers and senior decision makers.
  • Strong communication skills, openness to feedback.

Responsibilities

  • Build and extend Transluce’s AI evaluation methods for measuring evolving AI model behaviors.
  • Prototype and run behavioral evaluations for policy and oversight needs.
  • Execute on contracts with government evaluators and build evaluations for harmful manipulation.
  • Design and run privileged-access evaluations with frontier labs.
  • Adapt behavioral evaluation pipelines to contexts like mental health and persuasion.

Skills

Experience designing and running AI evaluations
Strong engineering instincts
Experience in customer-facing roles
Ability to balance needs of stakeholders
Strong communication skills

Job description

Salary range: $310,000 - $500,000/year + benefits

Description

Transluce is a fast-moving nonprofit research lab building the public tech stack for scalable AI evaluation and oversight. We specialize in behavioral evaluations of frontier AI systems, assessing how models actually behave in deployment, not just how they perform on benchmarks. We are an independent non-profit with a mission to steer the development of AI for the public good.

About the role

We're looking for an engineer to work on measuring and shaping AI model behaviors, someone who thrives on turning hard questions into evidence fast. Think of this as a forward‑deployed engineering role working directly with policymakers, civil society partners, and frontier labs to rapidly answer key questions about why AI systems act the way they do, and when and why they fail.

You'll build relationships with external domain experts, adapt our methods to new contexts, and help ensure our work is both technically credible and immediately useful to the people making consequential AI governance decisions.

This is a high‑autonomy role with direct exposure to senior stakeholders and a clear line of sight from your work to real‑world impact.

Core responsibility

Build and extend Transluce’s AI evaluation methods for measuring important evolving AI model behaviors. This includes:

  • Scope, prototype, and run behavioral evaluations in response to emerging policy and oversight needs, including rapid‑turnaround work for government and civil society partners.
  • Execute on Transluce’s contracts with government evaluators, including building evaluations for harmful manipulation with the EU AI Office.
  • Design and run privileged‑access evaluations and external oversight exercises with frontier labs.
  • Work with civil society organizations and domain experts to adapt our behavioral evaluation pipelines to their contexts (e.g., mental health, persuasion, evaluation awareness).
Qualities of a strong candidate
  • Hands‑on experience designing and running AI evaluations, particularly behavioral or interactive evaluations (multi‑turn, agentic, or red‑team contexts)
  • Strong engineering instincts and good judgment about when "good enough to ship" is actually good enough.
  • Experience in customer‑facing, consulting, or forward‑deployed roles translating ambiguous stakeholder needs into concrete deliverables.
  • Experience running evaluations at scale or in a production context.
  • Ability to understand and balance between the needs of AI researchers and domain experts, as well as between researchers and senior decision makers.
  • Strong communication skills, low ego, openness to giving and receiving feedback.

We are located in San Francisco and enthusiastic to work together in‑person. We are open to sponsoring international visas.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Behavior Researcher - Child Safety and Mental Health
AI Behavior Researcher - Child Safety and Mental Health

Transluce • San Francisco (CA)

On-site
USD 250,000 - 450,000
Research Engineer - Scalable Interpretability
Research Engineer - Scalable Interpretability

Transluce • San Francisco (CA)

On-site
USD 250,000 - 500,000
Forward‑Deployed AI Behavior Engineer
Forward‑Deployed AI Behavior Engineer

Transluce • San Francisco (CA)

On-site
USD 310,000 - 500,000
AI Behavior Researcher: Safeguarding Kids & Mental Health
AI Behavior Researcher: Safeguarding Kids & Mental Health

Transluce • San Francisco (CA)

On-site
USD 250,000 - 450,000
AI Research scientist - Evals
AI Research scientist - Evals

Cerebro • San Francisco (CA)

On-site
USD 165,000 - 195,000
Equity
Relocation support
Health and dental insurance
+2
Research Engineer, Model Evaluations
Research Engineer, Model Evaluations

Anthropic • New York (NY), San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Generous vacation and parental leave
Flexible working hours
Lovely office space for collaboration
AI Evaluation Engineer
AI Evaluation Engineer

DeepRec.ai • Denver (CO)

Remote
USD 180,000
Research Engineer, Model Evaluations
Research Engineer, Model Evaluations

Menlo Ventures • New York (NY)

On-site
USD 500,000 - 850,000
Pre-training Distributed Systems Tech Lead / Manager
Pre-training Distributed Systems Tech Lead / Manager

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Equity donation matching
Generous vacation and parental leave
Flexible working hours
+1
Research Engineer / Research Scientist - Model Behavior at OpenAI San Francisco, CA
Research Engineer / Research Scientist - Model Behavior at OpenAI San Francisco, CA

OpenAI • San Francisco (CA)

On-site
USD 310,000 - 460,000
Equity offers