Research Engineer: Scalable AI Interpretability

Transluce

San Francisco (CA)

On-site

USD 250,000 - 500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Transluce, a non-profit research lab in San Francisco, seeks scientists and engineers to advance AI oversight tools. You will help develop interpretable assistants and evaluate model behaviors, ensuring industry standards. Ideal candidates have experience in fine-tuning language models, excellent communication skills, and a curiosity for machine learning.

We offer a competitive salary range of $250,000 - $500,000 annually, in a collaborative, in-person work environment, with visa sponsorship available for international talents.

Qualifications

  • Experience with fine-tuning language models, designing new architectures, and creating evaluations.
  • Reliable results from good experimental design.
  • Strong programming skills and ability to navigate trade-offs between speed and maintainability.

Responsibilities

  • Help develop and train scalable interpretability assistants.
  • Create diverse evaluations that find undesirable model behaviors.
  • Scale up training and inference pipelines for large models.

Skills

Fine-tuning language models
Experimental design
Strong programming ability
Communication skills

Job description

Transluce, a non-profit research lab in San Francisco, seeks scientists and engineers to advance AI oversight tools. You will help develop interpretable assistants and evaluate model behaviors, ensuring industry standards. Ideal candidates have experience in fine-tuning language models, excellent communication skills, and a curiosity for machine learning.

We offer a competitive salary range of $250,000 - $500,000 annually, in a collaborative, in-person work environment, with visa sponsorship available for international talents.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer - Scalable Interpretability
Research Engineer - Scalable Interpretability

Transluce • San Francisco (CA)

On-site
USD 250,000 - 500,000
Research Engineer, Interpretability
Research Engineer, Interpretability

Anthropic • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Forward‑Deployed AI Behavior Engineer
Forward‑Deployed AI Behavior Engineer

Transluce • San Francisco (CA)

On-site
USD 310,000 - 500,000
Research Engineer - Mechanistic Interpretability
Research Engineer - Mechanistic Interpretability

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
AI Behavior Engineer
AI Behavior Engineer

Transluce • San Francisco (CA)

On-site
USD 310,000 - 500,000
Research Manager, Interpretability — Lead Safe AI Teams
Research Manager, Interpretability — Lead Safe AI Teams

Anthropic • San Francisco (CA)

On-site
USD 340,000 - 425,000
Interpretability Engineer for AI Safety & Research
Interpretability Engineer for AI Safety & Research

Anthropic • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Software Engineer, Infrastructure, Interpretability
Software Engineer, Infrastructure, Interpretability

United States Digital Space LLC • San Francisco (CA), New York (NY)

On-site
USD 320,000 - 485,000
AI Behavior Researcher: Safeguarding Kids & Mental Health
AI Behavior Researcher: Safeguarding Kids & Mental Health

Transluce • San Francisco (CA)

On-site
USD 250,000 - 450,000
Infrastructure Engineer, Interpretability & AI Safety
Infrastructure Engineer, Interpretability & AI Safety

Anthropic • San Francisco (CA)

On-site
USD 190,000 - 270,000