Founding Research Engineer — AI Safety & Frontier Models
Triage
San Francisco (CA)
On-site
USD 150,000 - 230,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Job summary
Triage is seeking a Founding Research Engineer in San Francisco to drive research for AI safety. You'll handle adversarial research, build training data infrastructure, and work on delivering effective detection systems. The role requires strong experience in LLM post-training and the ability to publish impactful research. With a competitive salary of $150K–$230K plus equity, this is an exciting opportunity to contribute to AI alignment and security.
Qualifications
Experience with reinforcement learning, DPO, instruction tuning, or adversarial training.
Ability to write coordinated disclosures and documentation.
Comfortable with quickly implementing new research.
Responsibilities
Run adversarial research on frontier model behavior.
Train and evaluate adversarially-trained models in production.
Drive specific research bets on customer-specific safety policies.
Publish coordinated disclosures and technical write-ups.
Skills
LLM post-training work
Technical opinions about misalignment
Strong writing skills
Ability to implement recent research quickly
Education
Bachelor’s, Master’s, or PhD in ML research or applied AI
Job description
Triage is seeking a Founding Research Engineer in San Francisco to drive research for AI safety. You'll handle adversarial research, build training data infrastructure, and work on delivering effective detection systems. The role requires strong experience in LLM post-training and the ability to publish impactful research. With a competitive salary of $150K–$230K plus equity, this is an exciting opportunity to contribute to AI alignment and security.