Founding Research Engineer — AI Safety & Frontier Models

Triage

San Francisco (CA)

On-site

USD 150,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Triage is seeking a Founding Research Engineer in San Francisco to drive research for AI safety. You'll handle adversarial research, build training data infrastructure, and work on delivering effective detection systems. The role requires strong experience in LLM post-training and the ability to publish impactful research. With a competitive salary of $150K–$230K plus equity, this is an exciting opportunity to contribute to AI alignment and security.

Qualifications

  • Experience with reinforcement learning, DPO, instruction tuning, or adversarial training.
  • Ability to write coordinated disclosures and documentation.
  • Comfortable with quickly implementing new research.

Responsibilities

  • Run adversarial research on frontier model behavior.
  • Train and evaluate adversarially-trained models in production.
  • Drive specific research bets on customer-specific safety policies.
  • Publish coordinated disclosures and technical write-ups.

Skills

LLM post-training work
Technical opinions about misalignment
Strong writing skills
Ability to implement recent research quickly

Education

Bachelor’s, Master’s, or PhD in ML research or applied AI

Job description

Triage is seeking a Founding Research Engineer in San Francisco to drive research for AI safety. You'll handle adversarial research, build training data infrastructure, and work on delivering effective detection systems. The role requires strong experience in LLM post-training and the ability to publish impactful research. With a competitive salary of $150K–$230K plus equity, this is an exciting opportunity to contribute to AI alignment and security.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist — Safe, Steerable AI & LLMs
Research Scientist — Safe, Steerable AI & LLMs

Menlo Ventures • San Francisco (CA)

Hybrid
USD 340,000 - 425,000
Competitive compensation and benefits
Generous vacation and parental leave
Flexible working hours
AI Safety & Research Engineer Fellow
AI Safety & Research Engineer Fellow

10a Labs • San Francisco (CA)

On-site
USD 80,000 - 120,000
Senior AI Safety Research Engineer
Senior AI Safety Research Engineer

Aisafety • Berkeley (CA)

Hybrid
USD 150,000 - 250,000
Catered lunch and dinner
Visa sponsorship for in-person employees
Work-related travel expenses covered
Senior Founding Recruiter for AI Safety & Research
Senior Founding Recruiter for AI Safety & Research

Aisafety • San Francisco (CA)

On-site
USD 120,000 - 280,000
Market competitive salary
Equity
Competitive benefits
AI Safety Research Engineer
AI Safety Research Engineer

Neura Market • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Frontier AI Safety Researcher
Frontier AI Safety Researcher

Triwill Group • San Francisco (CA)

On-site
USD 180,000 - 320,000
Research Lead, AI Safety & Impact Architect
Research Lead, AI Safety & Impact Architect

Aisafety • Berkeley (CA)

Hybrid
USD 170,000 - 270,000
Catered lunch and dinner
Visa sponsorship
Research Scientist — AI Safety & Alignment (Remote)
Research Scientist — AI Safety & Alignment (Remote)

Aisafety • Berkeley (CA)

Hybrid
USD 120,000 - 190,000
Catered lunch and dinner
Visa sponsorship for in-person employees
Reimbursement for work-related travel
AI Alignment Research Scientist — Safety & Reasoning
AI Alignment Research Scientist — Safety & Reasoning

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 150,000
AI Safety & Alignment Research Scientist
AI Safety & Alignment Research Scientist

Google DeepMind • Mountain View (CA)

On-site
USD 130,000 - 160,000