Frontier AI Safety Researcher

openai

General Trias

On-site

PHP 10,618,000 - 15,615,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

OpenAI is seeking exceptional researchers to push the frontier of safety mitigations for frontier AI models. You will derisk models by developing novel safety techniques drawn from interpretability, control, and alignment, contributing to a safer future of AGI.

This role requires strong technical depth and collaboration across misalignment, cybersecurity, and biology domains to build scalable, end-to-end safety stacks and robust safety systems.

Qualifications

  • 2+ years of experience in AI safety or related field.
  • Ph.D. or equivalent in computer science, machine learning, or related field.
  • Proficiency in Python or similar languages.
  • Experience in RLHF, interpretability, robustness, alignment, or control.

Responsibilities

  • Identify emerging AI safety risks and develop methodologies to mitigate them.
  • Build and refine evaluations to assess risk impact with domain experts as needed.
  • Set research directions to improve safety, alignment, and robustness of AI systems.
  • Contribute to best practices guidelines for AI safety for OpenAI and the industry.
  • Design and evaluate red-teaming pipelines to test end-to-end safety of deployed models.

Skills

Python
RLHF
Interpretability
Robustness
Control
Human-AI collaboration

Education

Ph.D. in Computer Science

Job description

OpenAI is seeking exceptional researchers to push the frontier of safety mitigations for frontier AI models. You will derisk models by developing novel safety techniques drawn from interpretability, control, and alignment, contributing to a safer future of AGI.

This role requires strong technical depth and collaboration across misalignment, cybersecurity, and biology domains to build scalable, end-to-end safety stacks and robust safety systems.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Researcher, Frontier Risk Mitigations
Researcher, Frontier Risk Mitigations

openai • General Trias

On-site
PHP 10,618,000 - 15,615,000
Researcher, Recursive Self-Improvement Safety
Researcher, Recursive Self-Improvement Safety

openai • General Trias

On-site
PHP 1,200,000 - 1,800,000
GenAI Security Researcher: AI Safety & Attack Research (Remote)
GenAI Security Researcher: AI Safety & Attack Research (Remote)

AIFT • Hinoba-an

Hybrid
PHP 4,719,764 - 5,899,705
Competitive compensation
Flexible working arrangements
Professional growth opportunities
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • Quezon City

On-site
PHP 600,000 - 1,200,000
Fullstack Engineer: Child Safety Intelligence & Tools
Fullstack Engineer: Child Safety Intelligence & Tools

openai • General Trias

On-site
PHP 8,745,000 - 12,492,000
Remote AI Safety & Policy Guardian
Remote AI Safety & Policy Guardian

TELUS Digital AI Data Solutions • Philippines

On-site
PHP 1,554,000 - 4,316,000
Sr. AI Security Engineer
Sr. AI Security Engineer

Backblaze • Mexico

On-site
PHP 3,683,241 - 5,524,861
Fullstack Software Engineer, Child Safety Tools & Systems
Fullstack Software Engineer, Child Safety Tools & Systems

openai • General Trias

On-site
PHP 8,745,000 - 12,492,000
Remote AI Safety & Adversarial Red Team Expert
Remote AI Safety & Adversarial Red Team Expert

Mercor • Philippines

On-site
PHP 4,316,000 - 6,782,000
Lead AI Research Scientist — Frontier ML & Optimization
Lead AI Research Scientist — Frontier ML & Optimization

Maya • Hinoba-an

On-site
PHP 1,800,000 - 3,500,000