Founding Research Engineer — AI Safety & Frontier Models

Triage

San Francisco (CA)

On-site

USD 150,000 - 230,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Triage is seeking a Founding Research Engineer in San Francisco to drive research for AI safety. You'll handle adversarial research, build training data infrastructure, and work on delivering effective detection systems. The role requires strong experience in LLM post-training and the ability to publish impactful research. With a competitive salary of $150K–$230K plus equity, this is an exciting opportunity to contribute to AI alignment and security.

Qualifications

  • Experience with reinforcement learning, DPO, instruction tuning, or adversarial training.
  • Ability to write coordinated disclosures and documentation.
  • Comfortable with quickly implementing new research.

Responsibilities

  • Run adversarial research on frontier model behavior.
  • Train and evaluate adversarially-trained models in production.
  • Drive specific research bets on customer-specific safety policies.
  • Publish coordinated disclosures and technical write-ups.

Skills

LLM post-training work
Technical opinions about misalignment
Strong writing skills
Ability to implement recent research quickly

Education

Bachelor’s, Master’s, or PhD in ML research or applied AI

Job description

Triage is seeking a Founding Research Engineer in San Francisco to drive research for AI safety. You'll handle adversarial research, build training data infrastructure, and work on delivering effective detection systems. The role requires strong experience in LLM post-training and the ability to publish impactful research. With a competitive salary of $150K–$230K plus equity, this is an exciting opportunity to contribute to AI alignment and security.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Scientist — Safe, Steerable AI & LLMs
Research Scientist — Safe, Steerable AI & LLMs

Anthropic • San Francisco (CA)

Hybrid
USD 340,000 - 425,000
Competitive compensation and benefits
Generous vacation and parental leave
Flexible working hours
Founding Research Engineer, AI Safety & Law Evaluations
Founding Research Engineer, AI Safety & Law Evaluations

Employia G&A Consulting GmbH • California

On-site
USD 185,000 - 350,000
Competitive equity
Founding team exposure
AI Safety & Research Engineer Fellow
AI Safety & Research Engineer Fellow

10a Labs • San Francisco (CA)

On-site
USD 80,000 - 120,000
Senior AI Safety Research Engineer
Senior AI Safety Research Engineer

Aisafety • Berkeley (CA)

Hybrid
USD 150,000 - 250,000
Catered lunch and dinner
Visa sponsorship for in-person employees
Work-related travel expenses covered
AI Safety Research Scientist
AI Safety Research Scientist

Center for AI Safety • San Francisco (CA)

On-site
USD 140,000 - 200,000
Health Insurance for you and dependets
401K with 4% matching
Unlimited PTO
+3
Frontier AI Safety Researcher (Hybrid — SF)
Frontier AI Safety Researcher (Hybrid — SF)

AI Chopping Block, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Hybrid work model
Relocation assistance
Research Scientist — AI Safety & Alignment (Remote)
Research Scientist — AI Safety & Alignment (Remote)

Aisafety • Berkeley (CA)

Hybrid
USD 120,000 - 190,000
Catered lunch and dinner
Visa sponsorship for in-person employees
Reimbursement for work-related travel
Research Engineer
Research Engineer

Confero • San Francisco (CA)

Hybrid
USD 200,000 - 300,000
Founding AI Engineer: Build Real-World AI Systems (Equity)
Founding AI Engineer: Build Real-World AI Systems (Equity)

Alpha Associates Recruitment • United States

Hybrid
USD 125,000 - 200,000
Founding-level equity
3% 401(k)
Unlimited PTO
+3
AI Alignment Research Scientist — Safety & Reasoning
AI Alignment Research Scientist — Safety & Reasoning

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 150,000