RL Research Scientist for Post-Training LLMs & Code Models

Advanced Micro Devices

Santa Clara, Northern (CA, KY)

Hybrid

USD 170,000 - 250,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Advanced Micro Devices (AMD) is hiring an AI Research Scientist specializing in reinforcement learning for post-training and interactive learning with large generative models used in demanding engineering tasks. You will invent and analyze RL methods, design reward models, and collaborate with infra and product teams to land methods that improve task success while maintaining stability and safety.

You will publish and ship, balancing theory with practical production-scale training, focusing on

Qualifications

  • Strong publication record in reinforcement learning or closely related ML areas.
  • Hands-on experience training RL or preference-optimized models at non-trivial scale (GPUs, distributed jobs).
  • Experience with LLM post-training, RLHF/RLAIF, or policy optimization for language or code agents.
  • Familiarity with compilers, kernels, EDA-style workflows, or large-scale codebases is a plus.

Responsibilities

  • Research and develop RL methods for post-training LLMs and code models on structured engineering tasks with verifiable or preference-based feedback.
  • Design reward models, curricula, and off-policy or on-policy training recipes suited to sparse, noisy, or expensive labels from experts and simulators.
  • Characterize failure modes (reward hacking, degenerate policies, instability) and propose mitigations grounded in experiments.
  • Collaborate with RL infra engineers to scale training; define interfaces for rollout generation, logging, and reproducibility.
  • Publish at top venues (e.g. NeurIPS, ICML, ICLR) and contribute internal technical leadership on the RL roadmap.

Skills

Reinforcement learning
LLMs
Policy optimization
Reward modeling

Education

PhD in Computer Science or related field

Job description

Advanced Micro Devices (AMD) is hiring an AI Research Scientist specializing in reinforcement learning for post-training and interactive learning with large generative models used in demanding engineering tasks. You will invent and analyze RL methods, design reward models, and collaborate with infra and product teams to land methods that improve task success while maintaining stability and safety.

You will publish and ship, balancing theory with practical production-scale training, focusing on

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RL Research Scientist - Post-Training on LLMs & Code Models
RL Research Scientist - Post-Training on LLMs & Code Models

AMD • Santa Clara (CA)

On-site
USD 150,000 - 230,000
Benefits at a glance
AI Research Scientist, Reinforcement Learning (LLM) and Post-Training
AI Research Scientist, Reinforcement Learning (LLM) and Post-Training

AMD • Santa Clara (CA)

On-site
USD 150,000 - 230,000
Benefits at a glance
AI Research Scientist, Reinforcement Learning (LLM) and Post-Training
AI Research Scientist, Reinforcement Learning (LLM) and Post-Training

Advanced Micro Devices • Santa Clara (CA), Northern (KY)

Hybrid
USD 170,000 - 250,000
Lead RL Infra Engineer - Scalable GPU Training Platforms
Lead RL Infra Engineer - Scalable GPU Training Platforms

Advanced Micro Devices • Santa Clara (CA)

On-site
USD 120,000 - 170,000
RL & Inference ML Systems Engineer for Engineering AI
RL & Inference ML Systems Engineer for Engineering AI

AMD • Santa Clara (CA)

On-site
USD 180,000 - 250,000
AMD benefits
Lead AI Research Scientist, Recursive Self Improvement, AI Safety and Reinforcement Learning
Lead AI Research Scientist, Recursive Self Improvement, AI Safety and Reinforcement Learning

AMD • Santa Clara (CA)

On-site
USD 120,000 - 180,000
ML Systems Engineer for RL & Inference Infrastructure
ML Systems Engineer for RL & Inference Infrastructure

Advanced Micro Devices • Santa Clara (CA)

Hybrid
USD 160,000 - 210,000
AMD benefits
ML Systems Research Engineer, RL / Inference / Agent Systems
ML Systems Research Engineer, RL / Inference / Agent Systems

Advanced Micro Devices • Santa Clara (CA)

Hybrid
USD 160,000 - 210,000
AMD benefits
Lead AI Research Scientist, Hardware AI Systems
Lead AI Research Scientist, Hardware AI Systems

Advanced Micro Devices • Santa Clara (CA)

On-site
USD 130,000 - 160,000
Comprehensive benefits package
Diversity and inclusion initiatives
ML Systems Research Engineer, RL / Inference / Agent Systems
ML Systems Research Engineer, RL / Inference / Agent Systems

AMD • Santa Clara (CA)

On-site
USD 180,000 - 250,000
AMD benefits