Research Engineer / Scientist, Alignment Science

Safetytalent

San Francisco (CA)

On-site

USD 100,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Safetytalent is seeking a Research Engineer / Scientist focused on Alignment Science in San Francisco, California. The role involves building and running machine learning experiments, contributing to research on AI safety, and collaborating on safety techniques.

The ideal candidate will work on designing evaluations to test models and contribute to AI safety efforts, including publications. This position emphasizes collaboration and innovation in understanding AI behavior.

Responsibilities

  • Build and run machine learning experiments to understand AI behavior.
  • Contribute to exploratory research on AI safety.
  • Collaborate on safety techniques testing projects.
  • Design evaluations to test models' reasoning abilities.
  • Contribute to research papers and blog posts.

Job description

Research Engineer / Scientist, Alignment Science

Anthropic

Responsibilities
  • Build and run machine learning experiments to understand and steer the behavior of powerful AI systems, focusing on making AI helpful, honest, and harmless.
  • Contribute to exploratory research on AI safety, particularly addressing risks from powerful future systems.
  • Collaborate with teams on projects such as testing safety techniques.
  • Design and implement evaluations to test models' reasoning abilities in safety-relevant contexts.
  • Contribute to research papers, blog posts, and experiments that feed into Anthropic's AI safety efforts.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist, AI Alignment & Safety
Research Scientist, AI Alignment & Safety

AI Breaking Wire • San Francisco (CA)

On-site
USD 300,000 - 450,000
Competitive salary and equity packages
Comprehensive health, dental, and visa
Flexible working arrangements and PTO
+1
Research Engineer: AI Safety & Alignment
Research Engineer: AI Safety & Alignment

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
AI Safety Research Engineer
AI Safety Research Engineer

Neura Market • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Research Engineer / Scientist, Alignment
Research Engineer / Scientist, Alignment

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
Research Engineer / Scientist, Alignment Anthropic San Francisco, CA
Research Engineer / Scientist, Alignment Anthropic San Francisco, CA

Neura Market • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Research Scientist, Alignment
Research Scientist, Alignment

AI Breaking Wire • San Francisco (CA)

On-site
USD 300,000 - 450,000
Competitive salary and equity packages
Comprehensive health, dental, and visa
Flexible working arrangements and PTO
+1
Research Engineer, AI Safety & Alignment
Research Engineer, AI Safety & Alignment

character • Redwood City (CA)

On-site
USD 120,000 - 150,000
Researcher, Alignment
Researcher, Alignment

OpenAI • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
AI Alignment Research Scientist — Safety & Reasoning
AI Alignment Research Scientist — Safety & Reasoning

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 150,000
Senior ML Engineer: AI Safety & Alignment (RLHF)
Senior ML Engineer: AI Safety & Alignment (RLHF)

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Top-tier salary and equity grants
Comprehensive medical, dental, and eye