Research Engineer – RL Infrastructure & Agent Environments

MaxIT Consulting - Max Corporate Group

San Francisco (CA)

On-site

USD 150,000 - 190,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

MaxIT Consulting - Max Corporate Group in San Francisco is seeking a Research Engineer – RL Infrastructure & Agent Environments to build the environments, evaluation systems, and supporting infrastructure used to train and assess long-horizon enterprise AI agents.

You will work on the engineering and research problems behind realistic agent environments, post-training systems, and reliable evaluation of complex multi-step workflows.

Qualifications

  • Hands-on experience with AI environments, evaluations, reinforcement learning infrastructure, or related agent-training systems.
  • Strong software engineering fundamentals.
  • Demonstrated ability to build and ship technical infrastructure.
  • Understanding of evaluation methodology, reward design, graders, and agent trajectories.
  • Ability to work across languages and technology stacks based on system requirements.

Responsibilities

  • Design evaluation environments for long-horizon enterprise agent workflows.
  • Define tasks, state, tools, graders, and reward signals used to evaluate and improve agents.
  • Build high-fidelity representations of complex enterprise software environments.
  • Develop infrastructure for rollouts, orchestration, trajectory inspection, and grader pipelines.
  • Measure both correctness and efficiency across multi-step agent behavior.
  • Investigate evaluation failures, reward-quality issues, and agent behavior.
  • Build production-quality systems rather than notebook-only research prototypes.

Skills

AI environments
Reinforcement learning
Infrastructure
Software engineering
Evaluation methodology
Reward design
Agent trajectories
Multilanguage stacks

Job description

San Francisco, California | Primarily On-site

We are seeking an Research Engineer – RL Infrastructure & Agent Environments to build the environments, evaluation systems, and supporting infrastructure used to train and assess long-horizon enterprise AI agents.

The Opportunity

You will work on the engineering and research problems behind realistic agent environments, post-training systems, and reliable evaluation of complex multi-step workflows.

Key Responsibilities
  • Design evaluation environments for long-horizon enterprise agent workflows.
  • Define tasks, state, tools, graders, and reward signals used to evaluate and improve agents.
  • Build high-fidelity representations of complex enterprise software environments.
  • Develop infrastructure for rollouts, orchestration, trajectory inspection, and grader pipelines.
  • Measure both correctness and efficiency across multi-step agent behavior.
  • Investigate evaluation failures, reward-quality issues, and agent behavior.
  • Build production-quality systems rather than notebook-only research prototypes.
Required Qualifications
  • Hands-on experience with AI environments, evaluations, reinforcement learning infrastructure, or related agent-training systems.
  • Strong software engineering fundamentals.
  • Demonstrated ability to build and ship technical infrastructure.
  • Understanding of evaluation methodology, reward design, graders, and agent trajectories.
  • Ability to work across languages and technology stacks based on system requirements.
Candidate Profile

A PhD is not required. Strong engineering and shipped environment or evaluation systems are more important than academic credentials or publication history.

Seniority

The opportunity is open to exceptional new graduates, early-career engineers, and experienced senior candidates. Selection is based primarily on engineering strength and relevant technical work.

Work Arrangement

The role is anchored in San Francisco with a strong preference for in-person collaboration. Limited flexibility may be considered case by case for exceptional candidates.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Evaluation Engineer – Reinforcement Learning & Agents
AI Evaluation Engineer – Reinforcement Learning & Agents

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 140,000 - 210,000
Research Engineer, RL Environments and Infrastructure
Research Engineer, RL Environments and Infrastructure

Hyphen Connect • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Production-Grade RL Environments Engineer
Production-Grade RL Environments Engineer

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 150,000 - 190,000
Forward Deployed AI Engineer – Agents & Research Systems
Forward Deployed AI Engineer – Agents & Research Systems

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 140,000 - 210,000
Full Stack Software Engineer - RL environments
Full Stack Software Engineer - RL environments

Acceler8 Talent • San Francisco (CA)

On-site
USD 120,000 - 180,000
RL Environment Software Engineer
RL Environment Software Engineer

talentpluto • San Francisco (CA)

On-site
USD 180,000 - 220,000
AI Evaluation Engineer: RL Environments & Agents
AI Evaluation Engineer: RL Environments & Agents

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 140,000 - 210,000
Software Engineer
Software Engineer

Acceler8 Talent • San Francisco (CA)

On-site
USD 225,000 - 275,000
Equity
Research Engineer
Research Engineer

Bespoke Labs • Mountain View (CA)

On-site
USD 120,000 - 140,000
Health coverage
Opportunity to work with leading AI labs
Competitive salary and equity
Research Engineer, Infrastructure, RL Systems
Research Engineer, Infrastructure, RL Systems

Thinkingmachines • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1