Research Engineer, AI Safety & Evaluations

TTN Talent

San Francisco (CA)

On-site

USD 200,000 - 400,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

TTN Talent in San Francisco is seeking a hands-on software engineer passionate about AI alignment, safety, and security. You will own technical projects end to end, from idea through implementation, testing and refinement.

You will build environments and evaluation tasks to identify misaligned or unreliable model behavior, and improve tooling to isolate, measure, and refine those behaviors. Salary on offer is $200,000 to $400,000.

Qualifications

  • Strong hands-on software engineering experience with focus on AI alignment, safety or security.
  • Experience with Python, agent tooling, evals, RL environments or similar technical infrastructure.
  • Ability to independently own technical projects end to end, from idea through implementation, testing and refinement.

Responsibilities

  • Build environments and evaluation tasks to identify misaligned or unreliable model behaviour, and improve software used to isolate, measure and refine those behaviours.

Job description

  • Do you have strong hands-on software engineering experience and a serious interest in AI alignment, AI safety or security?
  • Have you built, evaluated or directed LLM agents while they perform complex technical tasks?
  • Can you independently own technical projects end to end, from initial idea through implementation, testing and refinement?
  • Do you have experience with Python, agent tooling, evals, RL environments or similar technical infrastructure?

If you answer yes to all of the above, then this confidential role could be the one for you.

This is a rare opportunity to join a small, fast-growing AI company working on technical alignment and evaluation problems for frontier models. The successful candidate will independently build environments and evaluation tasks that help identify misaligned or unreliable model behaviour, while improving the software used to isolate, measure and refine those behaviours.

The role is based in San Francisco. Salary on offer is $200,000 to $400,000

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer — AI Alignment & Evaluation
Research Engineer — AI Alignment & Evaluation

W3 Sourcing • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Research Engineer
Research Engineer

Confero • San Francisco (CA)

Hybrid
USD 200,000 - 300,000
Researcher, Agent Safety, Training and Evaluations
Researcher, Agent Safety, Training and Evaluations

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 500,000
Relocation assistance
Hybrid work model
AI Alignment Research Engineer — Evaluation & Safety
AI Alignment Research Engineer — Evaluation & Safety

W3 Sourcing • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Researcher, Agent Safety, Training and Evaluations
Researcher, Agent Safety, Training and Evaluations

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Hybrid work model
Relocation assistance
AI Safety & Evaluations Engineer (End-to-End)
AI Safety & Evaluations Engineer (End-to-End)

TTN Talent • San Francisco (CA)

On-site
USD 200,000 - 400,000
AI Safety Research Engineer - RL Environments & Alignment
AI Safety Research Engineer - RL Environments & Alignment

Confero • San Francisco (CA)

Hybrid
USD 200,000 - 300,000
Machine Learning Engineer, Safety
Machine Learning Engineer, Safety

Harrison Clarke • San Francisco (CA)

Hybrid
USD 190,000 - 275,000
Founding Software Engineer – AI
Founding Software Engineer – AI

Socket.dev • San Francisco (CA)

On-site
USD 185,000 - 250,000
Equity opportunity
Leadership growth potential
Founding Software Engineer – AI
Founding Software Engineer – AI

Premier Global Links • San Francisco (CA)

On-site
USD 185,000 - 250,000
Equity