AI Safety Researcher — Investigate Model Behavior & Impact

Typesafe AI, Inc.

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

7 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

TypeSafe in San Francisco seeks a safety researcher to build evals, investigate concerning model behavior, and contribute to rigorous safety discussions. You will document observations from project start to formal records and draft internal communications.

You will work with highly talented researchers focused on AI safety, alignment, and interpretability, collaborating across teams to strengthen safety review processes.

Qualifications

  • Strong research background in AI safety, alignment, interpretability, or model evaluation.
  • Experience raising concerns calmly, then in writing.
  • The ability to say “I can't discuss specifics” and still communicate clearly.
  • A sincere belief that you can change things from the inside during early weeks.

Responsibilities

  • Build evals and investigate model behavior that concerns you.
  • Ask “should we?” in meetings discussing timelines and priorities.
  • Keep a document that starts as “Initial observations” and ends as “For the record.”
  • Write a departure memo that prompts reflection while maintaining discretion.

Skills

AI safety research
Model evaluation
Interpretability
Calmly raising concerns

Job description

TypeSafe in San Francisco seeks a safety researcher to build evals, investigate concerning model behavior, and contribute to rigorous safety discussions. You will document observations from project start to formal records and draft internal communications.

You will work with highly talented researchers focused on AI safety, alignment, and interpretability, collaborating across teams to strengthen safety review processes.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Safety Researcher (6-Week) — Evaluate & Document
AI Safety Researcher (6-Week) — Evaluate & Document

TypeSafe AI • San Francisco (CA)

On-site
USD 150,000 - 190,000
AI Safety Researcher: Alignment, Evaluation & Red Teaming
AI Safety Researcher: Alignment, Evaluation & Red Teaming

Thinking Machines Lab Inc. • San Francisco (CA)

On-site
USD 350,000 - 475,000
Visa sponsorship
Unlimited PTO
Parental leave
+2
Frontier-Model Safety Researcher & Evaluations
Frontier-Model Safety Researcher & Evaluations

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 500,000
Relocation assistance
Hybrid work model
AI Safety Engineer — Research to Production Leader
AI Safety Engineer — Research to Production Leader

Abundant • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 450,000
Health insurance
Dental insurance
Vision insurance
+1
Staff Software Engineer, AI Safety & Abuse Detection
Staff Software Engineer, AI Safety & Abuse Detection

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 485,000
AI Governance & Policy Research Lead
AI Governance & Policy Research Lead

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 130,000
Safety Systems Engineer: Build AI Safety Tools
Safety Systems Engineer: Build AI Safety Tools

OpenAI • San Francisco (CA)

On-site
USD 207,000 - 385,000
Research, Safety
Research, Safety

Thinking Machines Lab Inc. • San Francisco (CA)

On-site
USD 350,000 - 475,000
Visa sponsorship
Unlimited PTO
Parental leave
+2
Researcher, Safety Training, National Security
Researcher, Safety Training, National Security

Precision Labs • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Senior AI Safety & Oversight Researcher
Senior AI Safety & Oversight Researcher

OpenAI • San Francisco (CA)

On-site
USD 295,000 - 445,000