Frontier AI Evaluation & Environments Researcher

OpenAI

California (MO)

On-site

USD 180,000 - 230,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

OpenAI is hiring a researcher for Frontier Evals & Environments to build north star model environments and guide research programs behind their frontier models. You will work with researchers, engineers, product, infrastructure, and safety teams to decide what goes into major model runs, measure success, and ship improvements for real-world use.

You will thrive if you have strong ML and systems fundamentals, hands-on experience with LLMs and RL, and enjoy turning vague problems into concrete

Qualifications

  • Strong technical fundamentals in machine learning, software engineering, systems, or statistics.
  • Hands-on experience with LLMs, RL, RLHF/RLAIF, post-training, evals, graders, synthetic data, or production ML systems.
  • Ability to design scalable evaluation environments and run rigorous experiments (high-agency role).

Responsibilities

  • Create ambitious RL environments to push frontier models and measure capabilities, skills, and behaviors.
  • Develop new methodologies for automatically exploring model behavior.
  • Dive into measurement science, including scalability, reliability, and variance of evaluations.
  • Help steer training for large-scale runs and ship improvements into products.

Skills

Machine Learning
Software Engineering
Systems
Statistics
Research

Tools

Python
ML Frameworks
Evaluation Tools

Job description

OpenAI is hiring a researcher for Frontier Evals & Environments to build north star model environments and guide research programs behind their frontier models. You will work with researchers, engineers, product, infrastructure, and safety teams to decide what goes into major model runs, measure success, and ship improvements for real-world use.

You will thrive if you have strong ML and systems fundamentals, hands-on experience with LLMs and RL, and enjoy turning vague problems into concrete

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Frontier AI Evaluation & Environments Engineer
Frontier AI Evaluation & Environments Engineer

OpenAI • California (MO)

On-site
USD 180,000 - 240,000
Frontier AI Evaluation Scientist
Frontier AI Evaluation Scientist

Neura Market • San Francisco (CA)

On-site
USD 200,000 - 250,000
Frontier AI Evaluation Engineer: RL Environments
Frontier AI Evaluation Engineer: RL Environments

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Frontier RL Evaluation Engineer
Frontier RL Evaluation Engineer

Key Talent Solutions • San Francisco (CA)

On-site
USD 180,000 - 350,000
Frontier AI Safety Researcher
Frontier AI Safety Researcher

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 190,000 - 230,000
Relocation assistance
Frontier Health AI Research Engineer
Frontier Health AI Research Engineer

OpenAI • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Relocation assistance
Hybrid work model
Research Engineer, Frontier Evals & Environments
Research Engineer, Frontier Evals & Environments

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Research Engineer, Frontier Evals & Environments
Research Engineer, Frontier Evals & Environments

OpenAI • Los Angeles (CA)

On-site
USD 100,000 - 150,000
Research Engineer, Frontier Evals & Environments
Research Engineer, Frontier Evals & Environments

OpenAI • California (MO)

On-site
USD 180,000 - 240,000
Frontier AI ML Engineer — Evaluation & Production Systems
Frontier AI ML Engineer — Evaluation & Production Systems

Obsidian • New York (NY)

Remote
USD 4,000 - 7,000