AI Safety Researcher (6-Week) — Evaluate & Document

TypeSafe AI

San Francisco (CA)

On-site

USD 150,000 - 190,000

Full time

7 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

TypeSafe AI in San Francisco seeks a safety researcher to explore model behavior, raise concerns thoughtfully, and document findings for internal review.

The role covers building evals, interpreting alignment and evaluation results, and communicating clearly when asked to discuss specifics. You’ll join a team of highly talented researchers, with an emphasis on responsible disclosure and internal change from within.

Qualifications

  • Strong research background in AI safety, alignment, interpretability, or model evaluation.
  • Experience raising concerns calmly, then in writing.
  • Ability to say “I can't discuss specifics” and still maintain credibility.
  • Genuine belief you can drive change from within the organization.

Responsibilities

  • Build evals and investigate model behavior that concerns you.
  • Ask “should we?” in meetings scheduled to discuss “how soon?”.
  • Keep a document that starts as “Initial observations” and ends as “For the record.”
  • Write a departure post that makes people concerned enough to share it, despite containing almost no information.

Skills

AI safety
Calmly raise concerns
Clear written communication
Internal change mindset

Job description

TypeSafe AI in San Francisco seeks a safety researcher to explore model behavior, raise concerns thoughtfully, and document findings for internal review.

The role covers building evals, interpreting alignment and evaluation results, and communicating clearly when asked to discuss specifics. You’ll join a team of highly talented researchers, with an emphasis on responsible disclosure and internal change from within.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Safety Researcher — Investigate Model Behavior & Impact
AI Safety Researcher — Investigate Model Behavior & Impact

Typesafe AI, Inc. • San Francisco (CA)

On-site
USD 180,000 - 240,000
Frontier-Model Safety Researcher & Evaluations
Frontier-Model Safety Researcher & Evaluations

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 500,000
Relocation assistance
Hybrid work model
AI Governance & Policy Research Lead
AI Governance & Policy Research Lead

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 130,000
AI Safety & Oversight Systems Researcher
AI Safety & Oversight Systems Researcher

OpenAI • San Francisco (CA)

Hybrid
USD 195,000 - 230,000
Relocation assistance
Hybrid work model
AI Safety Engineer — Research to Production Leader
AI Safety Engineer — Research to Production Leader

Abundant • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 450,000
Health insurance
Dental insurance
Vision insurance
+1
Senior AI Safety & Oversight Researcher
Senior AI Safety & Oversight Researcher

OpenAI • San Francisco (CA)

On-site
USD 295,000 - 445,000
Staff Software Engineer, AI Safety & Abuse Detection
Staff Software Engineer, AI Safety & Abuse Detection

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 485,000
AI Safety Researcher — Equity & Sabbatical Perks
AI Safety Researcher — Equity & Sabbatical Perks

Best AI Tools Wiki • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Equity stake
Education budget
Relocation support
+1
AI Safety Practitioner
AI Safety Practitioner

DigiNo • Northern (KY)

Hybrid
USD 90,000 - 150,000
Researcher, Agent Safety, Training and Evaluations
Researcher, Agent Safety, Training and Evaluations

AI Chopping Block • San Francisco (CA), Northern (KY)

On-site
USD 180,000 - 240,000
Hybrid work model
Relocation assistance