Researcher, AI Agent Safety & Controls

CHEManager International

San Francisco (CA)

Hybrid

USD 180,000 - 260,000

Full time

11 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Relocation assistance
Hybrid work model

Job summary

OpenAI seeks a safety-focused researcher/engineer in San Francisco to design, implement, and test system-level controls for increasingly capable AI agents in real-world environments. The role emphasizes building practical mitigations, rigorous threat models, and evaluation infrastructure, with a strong focus on safety and alignment.

The role supports a hybrid work model, collaborating with Codex teams and other security specialists to advance responsible deployment of AI systems.

Qualifications

  • Experience designing system-level controls for AI or software systems.
  • Ability to reason about isolation boundaries, permissions and attack surfaces.
  • Experience building evaluative infrastructure and experiments to test mitigations.

Responsibilities

  • Design, build, and evaluate system-level controls for agent actions in real environments.
  • Collaborate with Codex harness engineering to productionize AI controls.
  • Red-team end-to-end agentic systems to test data exfiltration and unsafe tool use.
  • Improve safety-productivity tradeoffs by reducing missed harms and latency.

Skills

System security
Threat modeling
Experimentation
AI alignment
Risk assessment

Tools

Sandboxing
Process isolation
Codex harness

Job description

OpenAI seeks a safety-focused researcher/engineer in San Francisco to design, implement, and test system-level controls for increasingly capable AI agents in real-world environments. The role emphasizes building practical mitigations, rigorous threat models, and evaluation infrastructure, with a strong focus on safety and alignment.

The role supports a hybrid work model, collaborating with Codex teams and other security specialists to advance responsible deployment of AI systems.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Safety & Oversight Systems Researcher
AI Safety & Oversight Systems Researcher

OpenAI • San Francisco (CA)

Hybrid
USD 195,000 - 230,000
Relocation assistance
Hybrid work model
Safety & Systems Researcher for Autonomous AI Agents
Safety & Systems Researcher for Autonomous AI Agents

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 230,000
Relocation assistance
Health insurance
Researcher, Agent Safety, Oversight and System Mitigations
Researcher, Agent Safety, Oversight and System Mitigations

CHEManager International • San Francisco (CA)

Hybrid
USD 180,000 - 260,000
Relocation assistance
Hybrid work model
Researcher, Agent Safety, Oversight and System Mitigations
Researcher, Agent Safety, Oversight and System Mitigations

AI Chopping Block • San Francisco (CA), Northern (KY)

On-site
USD 150,000 - 230,000
Relocation assistance
Health insurance
Researcher, Agent Safety, Oversight and System Mitigations
Researcher, Agent Safety, Oversight and System Mitigations

OpenAI • San Francisco (CA)

On-site
USD 195,000 - 230,000
Relocation assistance
Hybrid work model
Frontier-Model Safety Researcher & Evaluations
Frontier-Model Safety Researcher & Evaluations

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 500,000
Relocation assistance
Hybrid work model
Agent Safety Researcher: Training & Evaluation
Agent Safety Researcher: Training & Evaluation

CHEManager International • San Francisco (CA)

Hybrid
USD 190,000 - 240,000
Senior AI Safety & Oversight Researcher
Senior AI Safety & Oversight Researcher

OpenAI • San Francisco (CA)

On-site
USD 295,000 - 445,000
Frontier AI Safety Researcher (Hybrid — SF)
Frontier AI Safety Researcher (Hybrid — SF)

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Hybrid work model
Relocation assistance
Researcher, Agent Safety, Training and Evaluations
Researcher, Agent Safety, Training and Evaluations

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 500,000
Relocation assistance
Hybrid work model