Agentic Safety Policy Lead for Frontier AI

Showcify

United States

On-site

USD 180,000 - 240,000

Full time

9 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

OpenAI is seeking a researcher to address real-world risks from model misalignment as AI systems scale in autonomy and horizon length. The role focuses on translating vulnerabilities into concrete behavioral policies, evaluations, and safeguards.

Hybrid role based in San Francisco with relocation support. You will work across safety systems, research, engineering, security, and policy to improve frontier AI safety and deployment decisions.

Qualifications

  • Strong background in AI safety, privacy, security, or adjacent fields.
  • Demonstrated interest in AI alignment and understanding of misaligned model behavior.

Responsibilities

  • Identify vulnerabilities as models interact with tools, data, and external systems; translate into safeguards.
  • Develop threat models and empirical frameworks for harmful outcomes from misalignment.
  • Build frameworks to understand harmful outcomes from model misalignment.
  • Identify underlying behaviors and system conditions driving those outcomes.
  • Translate findings into policy frameworks, evaluation criteria, and safeguards.
  • Develop human data campaigns and gold sets for measurement and evaluation of risks.
  • Collaborate with research, engineering, security, product to balance safety, utility and risk.
  • Inform deployment decisions, system cards, safeguards reports, and OpenAI’s agentic safety approach.
  • Build monitoring approaches to detect regressions and emerging risks post-deployment.

Skills

AI safety
Privacy
Security
Adversarial mindset
AI alignment
Data analysis
Evaluation data
Cross-functional collaboration
Clear communication

Job description

OpenAI is seeking a researcher to address real-world risks from model misalignment as AI systems scale in autonomy and horizon length. The role focuses on translating vulnerabilities into concrete behavioral policies, evaluations, and safeguards.

Hybrid role based in San Francisco with relocation support. You will work across safety systems, research, engineering, security, and policy to improve frontier AI safety and deployment decisions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Frontier AI Safety Researcher (Hybrid — SF)
Frontier AI Safety Researcher (Hybrid — SF)

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Hybrid work model
Relocation assistance
Frontier AI Safety Researcher: Mitigations & Alignment
Frontier AI Safety Researcher: Mitigations & Alignment

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 280,000
Model Policy & Agentic Safety Lead
Model Policy & Agentic Safety Lead

OpenAI • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Relocation support
Hybrid work model
Frontier-Model Safety Researcher & Evaluations
Frontier-Model Safety Researcher & Evaluations

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 500,000
Relocation assistance
Hybrid work model
Policy Lead for Agentic AI Safety
Policy Lead for Agentic AI Safety

United States Digital Space LLC • United States

On-site
USD 140,000 - 180,000
Model Policy Manager, Agentic Safety
Model Policy Manager, Agentic Safety

Showcify • United States

Hybrid
USD 180,000 - 240,000
Frontier AI Safety Evaluator & Alignment Expert
Frontier AI Safety Evaluator & Alignment Expert

Obsidian • New York (NY)

On-site
USD 120,000 - 180,000
Researcher, Agent Safety, Training and Evaluations
Researcher, Agent Safety, Training and Evaluations

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Hybrid work model
Relocation assistance
Frontier AI Safety & Transparency Lead
Frontier AI Safety & Transparency Lead

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Relocation assistance
Frontier AI Safety Evaluator & Policy Expert
Frontier AI Safety Evaluator & Policy Expert

Obsidian • San Francisco (CA)

On-site
USD 140,000 - 210,000