Research Engineer: RL & Reasoning for Next-Gen LMs

Zyphra

San Francisco (CA)

On-site

USD 100,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Comprehensive medical, dental, vision, and FSA plans
Competitive compensation and 401(k) plan
Relocation and immigration support
In-office snacks and meals provided
Unlimited PTO and company holidays
Collaborative team environment

Job summary

A cutting-edge AI company based in San Francisco is seeking a Research Engineer specializing in Agency and Reasoning. The role focuses on performing research in reinforcement learning and applying innovative ideas to the next generation of their language models. Candidates should have a postgraduate degree in a scientific field and be proficient in PyTorch and Python. The company values creativity and provides a dynamic work environment with excellent benefits, including comprehensive health plans and unlimited PTO.

Qualifications

  • Experience with reinforcement learning in language model reasoning.
  • Proficient with language-model-supervised fine-tuning and preference-learning methods.
  • Intuitive ability to understand model behaviors.

Responsibilities

  • Contribute to the Agency and Reasoning Team.
  • Perform novel research in reinforcement learning.
  • Apply ideas to next generation of language models.

Skills

Reinforcement learning
Prototyping
Python
PyTorch
Data engineering
Collaborative skills

Education

Postgraduate degree in a scientific subject

Tools

PyTorch
Python

Job description

A cutting-edge AI company based in San Francisco is seeking a Research Engineer specializing in Agency and Reasoning. The role focuses on performing research in reinforcement learning and applying innovative ideas to the next generation of their language models. Candidates should have a postgraduate degree in a scientific field and be proficient in PyTorch and Python. The company values creativity and provides a dynamic work environment with excellent benefits, including comprehensive health plans and unlimited PTO.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer/Research Scientist, RL/Reasoning
Research Engineer/Research Scientist, RL/Reasoning

OpenAI • Los Angeles (CA)

Hybrid
USD 120,000 - 160,000
Relocation assistance
Hybrid work model
Commitment to accommodation for disabilities
Research Engineer: ML & RL for Safe, Scalable AI
Research Engineer: ML & RL for Safe, Scalable AI

Anthropic • New York (NY)

On-site
USD 280,000 - 425,000
RL & Reasoning Research Scientist — Hybrid SF
RL & Reasoning Research Scientist — Hybrid SF

OpenAI • San Francisco (CA)

Hybrid
USD 295,000 - 445,000
Agent ML Research Engineer – Post-Training RL, Equity
Agent ML Research Engineer – Post-Training RL, Equity

Scale AI • United States

Hybrid
USD 180,000 - 315,000
Lead Research Scientist, LLM Agents & Systems
Lead Research Scientist, LLM Agents & Systems

NeoCognition Inc. • Palo Alto (CA)

On-site
USD 120,000 - 150,000
Research Engineer/Research Scientist, RL/Reasoning
Research Engineer/Research Scientist, RL/Reasoning

Slope • San Francisco (CA)

On-site
USD 310,000 - 460,000
Research Engineer: Multimodal RLHF & Personalized AI
Research Engineer: Multimodal RLHF & Personalized AI

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 445,000
Remote LLM Engineer - Research & Deployment
Remote LLM Engineer - Research & Deployment

Fastino Labs • San Francisco (CA)

Hybrid
GBP 65,000 - 90,000
Research Engineer — Scalable ML & RL for Safe AI
Research Engineer — Scalable ML & RL for Safe AI

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Research Scientist - Agentic LLMs & Foundations
Research Scientist - Agentic LLMs & Foundations

Ifm Us • Sunnyvale (CA)

On-site
USD 120,000 - 190,000
Comprehensive medical, dental, and vis
Bonus
401K Plan
+4