Senior RL Engineer - LLM Post-Training & Enterprise Apps

OneForma

United States

Hybrid

USD 140,000 - 210,000

Full time

4 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Centific AI Research seeks an Applied RL Engineer to design RL environments simulating enterprise workflows and train intelligent agents. You’ll bridge RL research with production systems, delivering measurable improvements in AI agent performance and safe, compliant deployments.

Responsibilities include building digital twins, post-training LLM agents with RLHF and PPO, and creating end-to-end training data pipelines. Strong Python & Gymnasium experience required.

Qualifications

  • Deep RL expertise: 3+ years hands-on experience with environment design, reward engineering, policy optimization.

Responsibilities

  • Design and build custom RL environments (digital twins) simulating enterprise workflows: document processing, compliance, onboarding, support automation

Skills

Deep RL
LLM post-training
Production software
Agentic AI
Python

Tools

Gymnasium
RLlib
Stable Baselines
PyTorch
JAX
TensorFlow

Job description

Centific AI Research seeks an Applied RL Engineer to design RL environments simulating enterprise workflows and train intelligent agents. You’ll bridge RL research with production systems, delivering measurable improvements in AI agent performance and safe, compliant deployments.

Responsibilities include building digital twins, post-training LLM agents with RLHF and PPO, and creating end-to-end training data pipelines. Strong Python & Gymnasium experience required.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist, LLM Evaluation & Post-Training
Research Scientist, LLM Evaluation & Post-Training

OneForma • United States

Hybrid
USD 140,000 - 210,000
Principal Research Scientist: AI Systems & RL in Production
Principal Research Scientist: AI Systems & RL in Production

Centific • East Palo Alto (CA)

On-site
USD 250,000 - 300,000
Research Engineer: RL & Post-Training LLM Systems
Research Engineer: RL & Post-Training LLM Systems

Preference Model • San Francisco (CA)

On-site
USD 120,000 - 150,000
Competitive cash and equity compensation (>90th percentile)
Health, vision, dental benefits
401K match
+2
Senior Applied Reinforcement Learning Engineer
Senior Applied Reinforcement Learning Engineer

Centific Global Solutions, Inc. • United States

Hybrid
USD 150,000 - 300,000
Remote RL Engineer: Advanced Reinforcement Learning
Remote RL Engineer: Advanced Reinforcement Learning

Booz Allen Hamilton • United States

Hybrid
USD 99,000 - 225,000
Research Engineer
Research Engineer

Oho Group • San Francisco (CA)

On-site
USD 160,000 - 230,000
Remote RL Research Intern — Agentic AI & LLMs
Remote RL Research Intern — Agentic AI & LLMs

Centific Global Solutions, Inc. • United States

On-site
Competitive stipend
Mentorship from researchers
Access to modern GPU infrastructure
RL Research Scientist - Post-Training on LLMs & Code Models
RL Research Scientist - Post-Training on LLMs & Code Models

AMD • Santa Clara (CA)

On-site
USD 150,000 - 230,000
Benefits at a glance
Principal Research Scientist, Reinforcement Learning
Principal Research Scientist, Reinforcement Learning

Centific • East Palo Alto (CA)

On-site
USD 250,000 - 300,000
Post-Training AI Lead for Enterprise RL & Evaluation
Post-Training AI Lead for Enterprise RL & Evaluation

Reflection AI • New York (NY)

On-site
USD 180,000 - 240,000
Stock options
Health insurance
Meals provided
+4