AI Research Engineer: Post-Training & Agentic RL

Oho Group

San Francisco (CA)

On-site

USD 160,000 - 230,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Oho Group is seeking an AI Researcher to advance post-training methods and agentic RL for complex hardware-engineering workflows. The role emphasizes hands-on development and practical deployment of AI agents in real-world tasks.

You will implement RL environments, generate synthetic data, and build benchmarks while collaborating with hardware experts to translate research into production-ready tools.

Qualifications

  • Proven experience with LLM post-training and RL environments.
  • Evidence of both research and practical implementation of AI agents.
  • Experience using Cursor or Claude Code in development workflows.

Responsibilities

  • Develop post-training approaches for agents performing multi-step tasks.
  • Build RL environments with executable tasks and reliable reward signals.
  • Generate and filter synthetic data using agent feedback to guide training.
  • Design benchmarks testing tool use and generalisation.
  • Collaborate with hardware specialists to map workflows to learnable tasks.
  • Turn research into working code and contribute tools/infrastructure.

Skills

Software engineering fundamentals
Post-training ML / RL
AI agents
Tooling (Cursor / Claude Code)

Tools

Cursor
Claude Code

Job description

Oho Group is seeking an AI Researcher to advance post-training methods and agentic RL for complex hardware-engineering workflows. The role emphasizes hands-on development and practical deployment of AI agents in real-world tasks.

You will implement RL environments, generate synthetic data, and build benchmarks while collaborating with hardware experts to translate research into production-ready tools.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Engineer
Research Engineer

Oho Group • San Francisco (CA)

On-site
USD 160,000 - 230,000
Lead RL & Agentic AI Research—On-Device & Tools
Lead RL & Agentic AI Research—On-Device & Tools

Socket.dev • Cupertino (CA)

On-site
USD 220,000 - 320,000
Lead RL Researcher for Agentic AI & Environments
Lead RL Researcher for Agentic AI & Environments

Apple Inc. • Cupertino (CA), Northern (KY)

Hybrid
USD 216,000 - 394,000
Apple stock programs
Medical coverage
Retirement benefits
+1
Remote AI Research Engineer — Agentic Post-Training
Remote AI Research Engineer — Agentic Post-Training

Tether.io • Town of Italy (NY)

On-site
Staff AI Engineer - Post-Training & Agentic RL
Staff AI Engineer - Post-Training & Agentic RL

Hark • San Jose (CA)

On-site
USD 180,000 - 450,000
Applied Scientist: Agentic RL & Post-Training
Applied Scientist: Agentic RL & Post-Training

Vecna AI • Chicago (IL)

On-site
USD 140,000 - 210,000
AI Research Scientist, New Grad – Agents & Reinforcement Learning
AI Research Scientist, New Grad – Agents & Reinforcement Learning

The available sources do not contain information about the company name for rounx.com. • United States

On-site
USD 140,000 - 190,000
AIML - Machine Learning Research Lead, RL Agents, MLR
AIML - Machine Learning Research Lead, RL Agents, MLR

Socket.dev • Cupertino (CA)

On-site
USD 220,000 - 320,000
AI Research Scientist: Multi-Agent & RL for Enterprise
AI Research Scientist: Multi-Agent & RL for Enterprise

AKIVA AI, LLC • Leesburg (VA)

On-site
USD 150,000 - 200,000
Professional development
Veteran-owned company values
Impactful client engagements
+8
Senior Research Scientist, Long-Horizon AI & RL (Equity)
Senior Research Scientist, Long-Horizon AI & RL (Equity)

techire ai • San Francisco (CA)

On-site
USD 400,000 - 450,000