RL Environments Engineer

MaxIT Consulting - Max Corporate Group

San Francisco (CA)

On-site

USD 140,000 - 200,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

MaxIT Consulting - Max Corporate Group in San Francisco, CA is seeking an Agent Evaluation Infrastructure Engineer to build environments, evaluation systems, and the supporting infrastructure used to train and assess long-horizon enterprise AI agents.

You will work on the engineering and research problems behind realistic agent environments, post-training systems, and reliable evaluation of complex multi-step workflows.

Qualifications

  • Hands-on experience with AI environments, evaluations, or RL infrastructure.
  • Strong software engineering fundamentals and production-quality infrastructure.
  • Understanding of evaluation metrics, reward design, graders, and trajectories.
  • Ability to work across languages and stacks as required by the system.

Responsibilities

  • Design evaluation environments for long-horizon enterprise agent workflows.
  • Define tasks, state, tools, graders, and reward signals used to evaluate and improve agents.
  • Build high-fidelity representations of complex enterprise software environments.
  • Develop infrastructure for rollouts, orchestration, trajectory inspection, and grader pipelines.
  • Measure both correctness and efficiency across multi-step agent behavior.
  • Investigate evaluation failures, reward-quality issues, and agent behavior.
  • Build production-quality systems rather than notebook-only research prototypes.

Skills

AI Environments
RL Infrastructure
Software Engineering
Infra Development
Evaluation Methodology
Cross-language Skills

Job description

San Francisco, California | Primarily On-site

We are seeking an Agent Evaluation Infrastructure Engineerto build the environments, evaluation systems, and supporting infrastructure used to train and assess long-horizon enterprise AI agents.

The Opportunity

You will work on the engineering and research problems behind realistic agent environments, post-training systems, and reliable evaluation of complex multi-step workflows.

Key Responsibilities
  • Design evaluation environments for long-horizon enterprise agent workflows.
  • Define tasks, state, tools, graders, and reward signals used to evaluate and improve agents.
  • Build high-fidelity representations of complex enterprise software environments.
  • Develop infrastructure for rollouts, orchestration, trajectory inspection, and grader pipelines.
  • Measure both correctness and efficiency across multi-step agent behavior.
  • Investigate evaluation failures, reward-quality issues, and agent behavior.
  • Build production-quality systems rather than notebook-only research prototypes.
Required Qualifications
  • Hands-on experience with AI environments, evaluations, reinforcement learning infrastructure, or related agent-training systems.
  • Strong software engineering fundamentals.
  • Demonstrated ability to build and ship technical infrastructure.
  • Understanding of evaluation methodology, reward design, graders, and agent trajectories.
  • Ability to work across languages and technology stacks based on system requirements.
Candidate Profile

A PhD is not required. Strong engineering and shipped environment or evaluation systems are more important than academic credentials or publication history.

Seniority

The opportunity is open to exceptional new graduates, early-career engineers, and experienced senior candidates. Selection is based primarily on engineering strength and relevant technical work.

Work Arrangement

The role is anchored in San Francisco with a strong preference for in-person collaboration. Limited flexibility may be considered case by case for exceptional candidates.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Agent Evaluation Infrastructure Engineer
Agent Evaluation Infrastructure Engineer

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 140,000 - 200,000
Enterprise Agent Systems Engineer
Enterprise Agent Systems Engineer

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 150,000 - 190,000
SF RL Agent Evaluation Infrastructure Engineer
SF RL Agent Evaluation Infrastructure Engineer

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 140,000 - 200,000
Enterprise AI Evaluation Infrastructure Engineer
Enterprise AI Evaluation Infrastructure Engineer

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 140,000 - 200,000
Applied AI Deployment Engineer
Applied AI Deployment Engineer

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 140,000 - 170,000
Staff Software Engineer, Systems Infrastructure - Agent Evaluation
Staff Software Engineer, Systems Infrastructure - Agent Evaluation

Linkedin3 • Mountain View (CA)

Hybrid
USD 180,000 - 240,000
RL Environments Engineer
RL Environments Engineer

Bespoke-Labs • Mountain View (CA)

On-site
USD 250,000 - 300,000
Health, dental, and vision
401(k)
Daily onsite lunch
+3
RL Environments Engineer
RL Environments Engineer

Bespoke Labs Inc. • Mountain View (CA), Northern (KY)

Hybrid
USD 250,000 - 300,000
Health, dental, and vision coverage
401(k)
Daily onsite lunch provided
+2
AI Lab Tech Engineer
AI Lab Tech Engineer

Commergence • Colorado

Hybrid
USD 150,000 - 210,000
RL Environment Software Engineer
RL Environment Software Engineer

talentpluto • San Francisco (CA)

Hybrid
USD 180,000 - 220,000