Open-Model Safety & Eval Program Manager

Reflection AI Ltd

New York (NY)

On-site

USD 140,000 - 200,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Top-tier comp
Stock options
Health plan
Meals provided
Parental leave
Unlimited PTO
Visa sponsorship
Team events

Job summary

Reflection AI Ltd. is seeking a Research Program Manager to embed with research and infrastructure teams, accelerating frontier model development.

This foundational role will define evaluation frameworks, build safety operations, and connect evals to the model lifecycle in a fast-moving, high-stakes environment. You will drive scoping of eval infrastructure, engage with researchers and engineers, and establish processes that scale as reflection interfaces with the broader safety ecosystem.

Qualifications

  • 7+ years in technical program management, research operations, or ML engineering.
  • Experience standing up new functions or programs from scratch.
  • Familiar with model evaluation and AI safety landscapes.
  • Able to engage with researchers and engineers on model behavior and data pipelines.

Responsibilities

  • Build the foundational infrastructure for model evals and safety.
  • Stand up model safety operations with workflows and cadences.
  • Embed safety checkpoints into development processes across training stages.
  • Establish engagement with external safety ecosystem and partnerships.
  • Provide visibility through reporting on model safety status and risks.

Skills

Technical program management
Research operations
ML engineering

Job description

Reflection AI Ltd. is seeking a Research Program Manager to embed with research and infrastructure teams, accelerating frontier model development.

This foundational role will define evaluation frameworks, build safety operations, and connect evals to the model lifecycle in a fast-moving, high-stakes environment. You will drive scoping of eval infrastructure, engage with researchers and engineers, and establish processes that scale as reflection interfaces with the broader safety ecosystem.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Program Manager — Model Safety & Eval
Research Program Manager — Model Safety & Eval

Reflection • New York (NY)

On-site
USD 120,000 - 150,000
Top-tier compensation
Comprehensive health benefits
Fully paid parental leave
+2
Research Program Manager - Model Evals and Safety
Research Program Manager - Model Evals and Safety

Reflection • New York (NY)

On-site
USD 120,000 - 150,000
Top-tier compensation
Comprehensive health benefits
Fully paid parental leave
+2
Research Program Manager - Model Evals and Safety
Research Program Manager - Model Evals and Safety

aijoblist • San Francisco (CA)

On-site
USD 130,000 - 180,000
Top-tier compensation
Comprehensive medical, dental, vision insurance
Paid parental leave
+2
Research Program Manager - Model Evals and Safety
Research Program Manager - Model Evals and Safety

Reflection AI Ltd • New York (NY)

On-site
USD 140,000 - 200,000
Top-tier comp
Stock options
Health plan
+5
Strategic Research Programs Lead in AI Infrastructure
Strategic Research Programs Lead in AI Infrastructure

Reflection AI Ltd • New York (NY)

On-site
USD 180,000 - 240,000
Top-tier compensation
Stock options
Health & wellness
+3
Safety Evaluations Platform Engineer – Secure Infra & Data
Safety Evaluations Platform Engineer – Secure Infra & Data

B Capital • San Francisco (CA)

On-site
USD 180,000 - 240,000
Top-tier compensation
Stock options
Health & wellness
+4
Adversarial Model Research PM: Safety & Evaluation Lead
Adversarial Model Research PM: Safety & Evaluation Lead

OpenAI • San Francisco (CA)

Hybrid
USD 239,000 - 328,000
Relocation assistance
Hybrid work model
AI Research Programs Lead
AI Research Programs Lead

Reflection • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive health insurance
Paid parental leave
+2
AI Research Infrastructure Lead
AI Research Infrastructure Lead

Reflection • New York (NY)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive health insurance
Paid parental leave
+2
Staff ML Evaluation & Data Engineer
Staff ML Evaluation & Data Engineer

Reflection • New York (NY)

On-site
USD 190,000 - 270,000
Stock options
Health & wellness benefits
Daily meals in office
+3