Researcher for Recursive AI Safety & Alignment

openai

General Trias

On-site

PHP 1,200,000 - 1,800,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

OpenAI Preparedness is seeking strong technical executors to anticipate and mitigate future misalignment risks in AI development. This role spans pre-deployment risk assessment, control measures, RSI-relevant training interventions, and translating insights into formal institutional practices and external communications.

The work is urgent and fast-paced, involving design, testing, and deployment of scalable safety solutions, with collaboration across engineering and policy teams to guide

Qualifications

  • Exceptional technical executor.
  • Strong strategic and research taste.
  • Passionate about mitigating risks associated with recursive self-improvement.

Responsibilities

  • Carefully consider the problems OpenAI might face in the future and how to prepare for them.
  • Turn an open-ended objective like “prepare for future misalignment threats” into a concrete direction (e.g. “stress-test monitors for scheming”) and prioritize the work.
  • Execute quickly, build scrappy prototypes, and iteratively improve them into established safety components.
  • Secure buy-in from other staff at OpenAI when necessary and communicate your work clearly.
  • Collaborate with or manage other staff as needed to scale efforts quickly.

Job description

OpenAI Preparedness is seeking strong technical executors to anticipate and mitigate future misalignment risks in AI development. This role spans pre-deployment risk assessment, control measures, RSI-relevant training interventions, and translating insights into formal institutional practices and external communications.

The work is urgent and fast-paced, involving design, testing, and deployment of scalable safety solutions, with collaboration across engineering and policy teams to guide

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Researcher, Recursive Self-Improvement Safety
Researcher, Recursive Self-Improvement Safety

openai • General Trias

On-site
PHP 1,200,000 - 1,800,000
Researcher, Frontier Risk Mitigations
Researcher, Frontier Risk Mitigations

openai • General Trias

On-site
PHP 10,618,000 - 15,615,000
Frontier AI Safety Researcher
Frontier AI Safety Researcher

openai • General Trias

On-site
PHP 10,618,000 - 15,615,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • Quezon City

On-site
PHP 600,000 - 1,200,000
AI Safety Red Team Tester — Remote, Flexible Hours
AI Safety Red Team Tester — Remote, Flexible Hours

Alignerr Corp. • Philippines

Remote
PHP 1,294,000 - 3,450,000
Fullstack Software Engineer, Child Safety Tools & Systems
Fullstack Software Engineer, Child Safety Tools & Systems

openai • General Trias

On-site
PHP 8,745,000 - 12,492,000
Fullstack Engineer: Child Safety Intelligence & Tools
Fullstack Engineer: Child Safety Intelligence & Tools

openai • General Trias

On-site
PHP 8,745,000 - 12,492,000
Sr. AI Security Engineer
Sr. AI Security Engineer

Backblaze • Mexico

On-site
PHP 3,683,241 - 5,524,861
AI Red Team Tester
AI Red Team Tester

Alignerr Corp. • Philippines

Remote
PHP 1,294,000 - 3,450,000
GenAI Security Researcher: AI Safety & Attack Research (Remote)
GenAI Security Researcher: AI Safety & Attack Research (Remote)

AIFT • Hinoba-an

Hybrid
PHP 4,719,764 - 5,899,705
Competitive compensation
Flexible working arrangements
Professional growth opportunities