RL Research Scientist for LLM Post-Training & Code Models

AMD

Santa Clara (CA)

On-site

USD 180,000 - 260,000

Full time

26 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

AMD benefits

Job summary

AMD in Santa Clara, CA is seeking an AI Research Scientist, Reinforcement Learning (LLM) and Post-Training to push RL methods for large generative models used in engineering tasks. You will design reward models, run empirical studies, and partner with infra and product teams to land robust, scalable approaches.

You should have a PhD and a strong publication record, with hands-on experience training RL or preference-based models at scale using GPUs and distributed systems.

Qualifications

  • PhD in Computer Science, Machine Learning, or related field strongly preferred.
  • Strong publication record in reinforcement learning or closely related ML areas.
  • Hands-on experience training RL or preference‑optimized models at non-trivial scale (GPUs, distributed jobs).

Responsibilities

  • Research and develop RL methods for post‑training LLMs and code models on structured engineering tasks with verifiable or preference‑based feedback.
  • Design reward models, curricula, and off‑policy or on‑policy training recipes suited to sparse, noisy, or expensive labels from experts and simulators.
  • Characterize failure modes (reward hacking, degenerate policies, instability) and propose mitigations grounded in experiments.
  • Collaborate with RL infra engineers to scale training; define interfaces for rollout generation, logging, and reproducibility.
  • Publish at top venues (e.g. NeurIPS, ICML, ICLR) and contribute internal technical leadership on the RL roadmap.

Skills

RL theory
Publication record
GPUs
Distributed training

Education

PhD in CS/ML

Tools

LLMs
RLHF

Job description

AMD in Santa Clara, CA is seeking an AI Research Scientist, Reinforcement Learning (LLM) and Post-Training to push RL methods for large generative models used in engineering tasks. You will design reward models, run empirical studies, and partner with infra and product teams to land robust, scalable approaches.

You should have a PhD and a strong publication record, with hands-on experience training RL or preference-based models at scale using GPUs and distributed systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Research Scientist, Reinforcement Learning (LLM) and Post-Training
AI Research Scientist, Reinforcement Learning (LLM) and Post-Training

AMD • Santa Clara (CA)

On-site
USD 180,000 - 260,000
AMD benefits
ML Engineer — LLM Post-Training & RL Specialist
ML Engineer — LLM Post-Training & RL Specialist

NewsBreak • Mountain View (CA)

On-site
USD 130,000 - 160,000
Health, dental, and vision care
401(k) plan with company matching
Paid time off and holidays
Research Engineer: RL & Post-Training LLM Systems
Research Engineer: RL & Post-Training LLM Systems

Preference Model • San Francisco (CA)

On-site
USD 120,000 - 150,000
Competitive cash and equity compensation (>90th percentile)
Health, vision, dental benefits
401K match
+2
Lead RL Infrastructure Engineer — Scalable GPU Training
Lead RL Infrastructure Engineer — Scalable GPU Training

AMD • Santa Clara (CA)

On-site
USD 130,000 - 180,000
Competitive benefits package
Generative AI Research Scientist: LLM Post-Training
Generative AI Research Scientist: LLM Post-Training

Scale AI, Inc. • New York (NY), Northern (KY)

Hybrid
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
ML Systems Engineer for RL & Inference Infrastructure
ML Systems Engineer for RL & Inference Infrastructure

Advanced Micro Devices • Santa Clara (CA)

Hybrid
USD 160,000 - 210,000
AMD benefits
Applied AI Research Scientist — LLMs, LMMs & Diffusion
Applied AI Research Scientist — LLMs, LMMs & Diffusion

Socket.dev • San Jose (CA)

Hybrid
USD 150,000 - 200,000
Machine Learning Researcher – LLM
Machine Learning Researcher – LLM

Susquehanna International Group • Bala Cynwyd (PA)

On-site
USD 180,000 - 280,000
RL Post-Training Scientist for LLMs & Tooling
RL Post-Training Scientist for LLMs & Tooling

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 167,000 - 226,000
Health insurance
401(k)
Post-Training AI Research Engineer – RL & Agentic Infra
Post-Training AI Research Engineer – RL & Agentic Infra

Storm3 • San Francisco (CA)

On-site
USD 140,000 - 210,000
Medical Insurance
Dental Insurance
Vision Insurance
+2