AI Safety & Evaluations Engineer (End-to-End)

TTN Talent

San Francisco (CA)

On-site

USD 200,000 - 400,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

TTN Talent in San Francisco is seeking a hands-on software engineer passionate about AI alignment, safety, and security. You will own technical projects end to end, from idea through implementation, testing and refinement.

You will build environments and evaluation tasks to identify misaligned or unreliable model behavior, and improve tooling to isolate, measure, and refine those behaviors. Salary on offer is $200,000 to $400,000.

Qualifications

  • Strong hands-on software engineering experience with focus on AI alignment, safety or security.
  • Experience with Python, agent tooling, evals, RL environments or similar technical infrastructure.
  • Ability to independently own technical projects end to end, from idea through implementation, testing and refinement.

Responsibilities

  • Build environments and evaluation tasks to identify misaligned or unreliable model behaviour, and improve software used to isolate, measure and refine those behaviours.

Job description

TTN Talent in San Francisco is seeking a hands-on software engineer passionate about AI alignment, safety, and security. You will own technical projects end to end, from idea through implementation, testing and refinement.

You will build environments and evaluation tasks to identify misaligned or unreliable model behavior, and improve tooling to isolate, measure, and refine those behaviors. Salary on offer is $200,000 to $400,000.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer, AI Safety & Evaluations
Research Engineer, AI Safety & Evaluations

TTN Talent • San Francisco (CA)

On-site
USD 200,000 - 400,000
AI Alignment Research Engineer — Evaluation & Safety
AI Alignment Research Engineer — Evaluation & Safety

W3 Sourcing • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
AI Safety Strategy & Talent Lead
AI Safety Strategy & Talent Lead

Aisafety • San Francisco (CA)

On-site
USD 160,000 - 300,000
Unlimited PTO
Comprehensive health insurance
10% employer 401(k) contribution
AI Safety & Oversight Engineer
AI Safety & Oversight Engineer

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
AI Alignment Research Scientist — Safety & Reasoning
AI Alignment Research Scientist — Safety & Reasoning

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 150,000
Staff AI Safety Engineer - Red Team & Guardrails
Staff AI Safety Engineer - Red Team & Guardrails

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive medical, dental, vision, life, and disability insurance
Fully paid parental leave
+2
AI Safety Research Engineer - RL Environments & Alignment
AI Safety Research Engineer - RL Environments & Alignment

Confero • San Francisco (CA)

Hybrid
USD 200,000 - 300,000
Delivery Engineer: AI Safety & Enterprise Evaluations
Delivery Engineer: AI Safety & Enterprise Evaluations

Artificial Intelligence Underwriting Company • San Francisco (CA)

On-site
USD 180,000 - 230,000
Competitive salary
Equity
Relocation to San Francisco
+1
Safety Systems Engineer: Build AI Safety Tools
Safety Systems Engineer: Build AI Safety Tools

OpenAI • San Francisco (CA)

On-site
USD 207,000 - 385,000
Research Engineer: AI Safety & Alignment
Research Engineer: AI Safety & Alignment

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours