Research Scientist, AI Alignment & Reasoning (Remote)

AI Breaking Wire

San Francisco, Northern (CA, KY)

Hybrid

USD 300,000 - 450,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Equity packages
Health insurance
Remote-friendly culture
Parental leave
Wellness stipends
SF office spaces

Job summary

Anthropic, a safety-focused AI research company in San Francisco, seeks a Research Scientist to advance reasoning and alignment in large language models. You will design experiments, build RLHF and constitutional AI pipelines, and collaborate with engineers to scale training.

You will publish results and contribute to frontier AI understanding, while enjoying a competitive salary, equity, comprehensive benefits, and a remote-friendly office culture.

Qualifications

  • Ph.D. or equivalent practical experience in Machine Learning, Computer Science, Statistics, or a related technical field.
  • Demonstrated publication record in top-tier AI/ML conferences (NeurIPS, ICML, ICLR, ACL).
  • Strong programming skills in Python and familiarity with PyTorch or JAX.
  • Deep understanding of transformer architectures, reinforcement learning, and alignment methodologies.
  • Passion for AI safety and responsible development of advanced artificial intelligence.

Responsibilities

  • Design and conduct novel experiments to improve the reasoning capabilities and safety alignment of large language models.
  • Develop advanced reinforcement learning from human feedback (RLHF) and constitutional AI training pipelines.
  • Collaborate with infrastructure engineers to scale training runs across thousands of accelerators.
  • Publish research findings and contribute to the broader scientific understanding of frontier AI models.

Skills

Python
PyTorch
RLHF
Alignment
Research
LLM

Education

Ph.D. or equivalent practical experience in Machine Learning, Computer Science, Statistics, or a related technical field

Tools

JAX

Job description

Anthropic, a safety-focused AI research company in San Francisco, seeks a Research Scientist to advance reasoning and alignment in large language models. You will design experiments, build RLHF and constitutional AI pipelines, and collaborate with engineers to scale training.

You will publish results and contribute to frontier AI understanding, while enjoying a competitive salary, equity, comprehensive benefits, and a remote-friendly office culture.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist, Reasoning and Alignment
Research Scientist, Reasoning and Alignment

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 300,000 - 450,000
Equity packages
Health insurance
Remote-friendly culture
+3
Research Scientist, AI Alignment & Safety
Research Scientist, AI Alignment & Safety

AI Breaking Wire • San Francisco (CA)

On-site
USD 300,000 - 450,000
Competitive salary and equity packages
Comprehensive health, dental, and visa
Flexible working arrangements and PTO
+1
Research Scientist, Alignment
Research Scientist, Alignment

AI Breaking Wire • San Francisco (CA)

On-site
USD 300,000 - 450,000
Competitive salary and equity packages
Comprehensive health, dental, and visa
Flexible working arrangements and PTO
+1
Research Engineer: AI Safety & Alignment
Research Engineer: AI Safety & Alignment

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
RL & Reasoning Research Scientist — Hybrid SF
RL & Reasoning Research Scientist — Hybrid SF

OpenAI • San Francisco (CA)

Hybrid
USD 295,000 - 445,000
Senior ML Engineer - AI Alignment & RLHF (Remote)
Senior ML Engineer - AI Alignment & RLHF (Remote)

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Equity compensation
Competitive salary
Remote and hybrid options
+1
AI Alignment Research Scientist — Safety & Reasoning
AI Alignment Research Scientist — Safety & Reasoning

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 150,000
Senior ML Engineer: AI Safety & Alignment (RLHF)
Senior ML Engineer: AI Safety & Alignment (RLHF)

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Top-tier salary and equity grants
Comprehensive medical, dental, and eye
Research Engineer
Research Engineer

Confero • San Francisco (CA)

Hybrid
USD 200,000 - 300,000
Research Engineer/Research Scientist, RL/Reasoning
Research Engineer/Research Scientist, RL/Reasoning

OpenAI • Los Angeles (CA)

Hybrid
USD 120,000 - 160,000
Relocation assistance
Hybrid work model
Commitment to accommodation for disabilities