Remote RL Research Intern — Agentic AI & LLMs

Centific Global Solutions, Inc.

United States

On-site

USD 48,216 - 61,992

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive stipend
Mentorship from researchers
Access to modern GPU infrastructure

Job summary

Centific Global Solutions, Inc. is seeking a PhD Research Intern in Applied Reinforcement Learning to design and evaluate RL systems for agentic AI workflows. The role involves developing RL environments, reward models, and post-training pipelines for LLM-based agents.

The ideal candidate will have a PhD in computer science, machine learning, or a related field, strong Python and PyTorch skills, and a solid understanding of RL fundamentals. This position offers a competitive stipend and mentorship opportunities.

Qualifications

  • PhD candidate with research in reinforcement learning or agentic AI.
  • Strong skills in Python and PyTorch, with GPU training experience.
  • Understanding of reinforcement learning fundamentals.

Responsibilities

  • Design and evaluate reinforcement learning systems for agentic AI workflows.
  • Develop and evaluate RL environments for enterprise workflows.
  • Prototype an agentic system integrated with RL training.

Skills

Python
PyTorch
Reinforcement Learning Fundamentals
GPU-based Training
Experimentation Practices

Education

PhD candidate in CS, ML, or related field

Tools

Gymnasium
RLlib
Stable Baselines

Job description

Centific Global Solutions, Inc. is seeking a PhD Research Intern in Applied Reinforcement Learning to design and evaluate RL systems for agentic AI workflows. The role involves developing RL environments, reward models, and post-training pipelines for LLM-based agents.

The ideal candidate will have a PhD in computer science, machine learning, or a related field, strong Python and PyTorch skills, and a solid understanding of RL fundamentals. This position offers a competitive stipend and mentorship opportunities.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Staff Scientist, Agentic AI & RL — Lead AI Systems
Senior Staff Scientist, Agentic AI & RL — Lead AI Systems

Centific Global Solutions, Inc. • Palo Alto (CA)

On-site
USD 150,000 - 200,000
Research Engineer: RL & Post-Training LLM Systems
Research Engineer: RL & Post-Training LLM Systems

Preference Model • San Francisco (CA)

On-site
USD 120,000 - 150,000
Competitive cash and equity compensation (>90th percentile)
Health, vision, dental benefits
401K match
+2
Research Intern — LLM Agents & RL Training
Research Intern — LLM Agents & RL Training

NewsBreak • Mountain View (CA)

On-site
Post-Training AI Research Engineer – RL & Agentic Infra
Post-Training AI Research Engineer – RL & Agentic Infra

Storm3 • San Francisco (CA)

On-site
USD 140,000 - 210,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
Research Engineer, Post-Training
Research Engineer, Post-Training

Storm3 • San Francisco (CA)

On-site
USD 140,000 - 210,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
AI Research Scientist, New Grad – Agents & Reinforcement Learning
AI Research Scientist, New Grad – Agents & Reinforcement Learning

The available sources do not contain information about the company name for rounx.com. • United States

On-site
USD 140,000 - 190,000
Research Intern
Research Intern

NeoCognition Inc. • Palo Alto (CA)

On-site
USD 60,000 - 80,000
Research Intern
Research Intern

NeoCognition • Palo Alto (CA)

On-site
Remote Senior RL Scientist — Agent Design & Training
Remote Senior RL Scientist — Agent Design & Training

Centific Global Solutions, Inc. • United States

On-site
USD 200,000 - 250,000
Research Intern: LLM Agents & Next-Gen Prototypes
Research Intern: LLM Agents & Next-Gen Prototypes

NeoCognition • Palo Alto (CA)

On-site