Model Policy, Frontier Cyber Risk

Slope

San Francisco (CA)

On-site

USD 207,000 - 295,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, and vision insurance
401(k) retirement plan with employer match
Paid parental leave
Flexible PTO
Mental health and wellness support
Annual learning and development stipend
Daily meals in the office

Job summary

Slope is looking for a cybersecurity policy specialist to define and implement policies for high-risk AI systems. This role involves designing model policies, building frameworks for implementation, and collaborating with engineering teams. Candidates should have strong technical expertise in cybersecurity and excellent communication skills. Located in San Francisco with a hybrid work model, the position offers competitive compensation ranging from $207,000 to $295,000, plus equity and generous benefits.

Qualifications

  • Experience in offensive security, defensive security, and vulnerability research.
  • Ability to distinguish between legitimate security and harmful activity.
  • Experience in building threat models for complex systems.

Responsibilities

  • Design model policies for cybersecurity and high-risk domains.
  • Build policy artifacts for implementation across various systems.
  • Analyze deployment data and improve policy quality over time.

Skills

Technical expertise in cybersecurity
Judgment on AI systems and cyber threats
Translating security expertise into policy
Strong communication skills

Job description

Location

San Francisco

Employment Type

Full time

Location Type

Hybrid

Department

Safety Systems

Compensation
  • San Francisco $207K – $295K • Offers Equity

The base pay offered may vary depending on multiple individualized factors, including market location, job‑related knowledge, skills, and experience. If the role is non‑exempt, overtime pay will be provided consistent with applicable laws. In addition to the salary range listed above, total compensation also includes generous equity, performance‑related bonus(s) for eligible employees, and the following benefits.

  • Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings Accounts
  • Pre‑tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)
  • 401(k) retirement plan with employer match
  • Paid parental leave (up to 24 weeks for birth parents and 20 weeks for non‑birthing parents), plus paid medical and caregiver leave (up to 8 weeks)
  • Paid time off: flexible PTO for exempt employees and up to 15 days annually for non‑exempt employees
  • 13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time (1 hour per 30 hours worked, or more, as required by applicable state or local law)
  • Mental health and wellness support
  • Employer‑paid basic life and disability coverage
  • Annual learning and development stipend to fuel your professional growth
  • Daily meals in our offices, and meal delivery credits as eligible
  • Relocation support for eligible employees
  • Additional taxable fringe benefits, such as charitable donation matching and wellness stipends, may also be provided.

More details about our benefits are available to candidates during the hiring process.

This role is at‑will and OpenAI reserves the right to modify base pay and other compensation components at any time based on individual performance, team or company results, or market conditions.

About the Team

Our Safety Systems team is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency.

Within Safety Systems, the Model Policy team aligns model behavior with desired human values and norms. We co‑design policy with models and for models by driving rapid policy taxonomy iteration based on data and defining evaluation criteria for foundational models’ ability to reason about safety.

About the Role

Frontier AI systems are rapidly expanding what is possible in cybersecurity and software engineering. These capabilities create major defensive opportunities, but they also raise serious dual‑use and misuse risks across areas such as malware development, exploit discovery, vulnerability chaining, credential abuse, cyber intrusion, and autonomous offensive operations.

In this role, you will help define how OpenAI’s models should behave in high‑risk cybersecurity contexts. You will develop policy frameworks, threat models, taxonomies, evaluations, and behavioral specifications that guide model behavior across training, deployment, and monitoring systems. This role sits at the intersection of cybersecurity, AI safety, threat modeling, evaluation science, and policy implementation.

You will work closely with research, engineering, safety training, preparedness, and product teams to build policies that are technically grounded, measurable, enforceable, and responsive to real‑world cyber risk.

Your Responsibilities
  • Design and maintain model policies for cybersecurity and frontier‑risk domains, especially dual‑use and high‑risk cyber capabilities.
  • Translate cybersecurity threat models into clear behavioral specifications, evaluation criteria, grading guidance, and system‑level mitigations.
  • Define practical boundaries between legitimate security research, defensive workflows, and assistance that could materially enable harmful activity.
  • Build policy artifacts that support implementation across training, evaluation, deployment, monitoring, and escalation systems.
  • Partner with safety researchers, engineers, and evaluation teams to operationalize policies into scalable model behavior and measurable safeguards.
  • Analyze red‑teaming results, deployment data, model failures, over‑refusals, and ambiguous edge cases to improve policy and evaluation quality over time.
  • Identify emerging cyber capability areas where advanced AI systems could lower barriers to misuse or increase operational capability for malicious actors.
  • Contribute to system cards, safety reports, policy documentation, and external communications on OpenAI’s approach to cyber risk mitigation.
We’re Seeking
  • Strong technical expertise in cybersecurity, such as offensive security, defensive security, vulnerability research, malware analysis, incident response, threat intelligence, application security, exploit development, infrastructure security, or cloud security.
  • Strong judgment about how AI systems may affect the cyber threat landscape, including dual‑use, autonomous, or agentic system risks.
  • Ability to distinguish between legitimate security use cases and assistance that could materially enable harmful cyber activity.
  • Experience building or applying threat models to complex technical systems, especially in adversarial or high‑risk environments.
  • Ability to translate technical security expertise into structured policy frameworks, evaluation criteria, operational guidance, and enforcement mechanisms.
  • Comfort using empirical evidence, including evaluations, red‑teaming results, deployment observations, and model failure modes, to inform policy decisions.
  • Strong systems thinking across policy, evaluations, classifiers, training, deployment safeguards, measurement, and monitoring.
  • Ability to work cross‑functionally with researchers, engineers, product teams, policy experts, and operational stakeholders.
  • Strong written communication skills, especially the ability to explain complex technical and security concepts clearly.
  • A pragmatic approach to safety: focused on reducing real‑world risk while preserving legitimate, beneficial, and defensive uses of AI.
Workplace & Location

This role is based in our San Francisco office. We do encourage you to apply even if you prefer a different work location as factors may change over time.

We offer relocation support to new employees, and we use a hybrid model: three days in the office per week with optional work from home on Thursdays and Fridays.

Our open‑plan offices have height‑adjustable desks, conference rooms, phone booths, well‑stocked kitchens full of snacks and drinks, three in‑house prepared meals daily, a private outdoor space for working in the sun or socializing, nap rooms, private bike storage, and more.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general‑purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

OpenAI Global Applicant Privacy Policy

Compensation Range: $207K - $295K

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Model Policy
Model Policy

Slope • San Francisco (CA)

On-site
USD 207,000 - 295,000
Medical, dental, and vision insurance
401(k) retirement plan with employer match
Paid parental leave
+3
Model Policy, Frontier Cyber Risk
Model Policy, Frontier Cyber Risk

OpenAI • Los Angeles (CA)

Hybrid
USD 207,000 - 295,000
Model Policy, Frontier Cyber Risk
Model Policy, Frontier Cyber Risk

OpenAI • San Francisco (CA)

Hybrid
USD 207,000 - 295,000
Model Policy
Model Policy

Neura Market • San Francisco (CA)

Hybrid
USD 180,000 - 280,000
Researcher, Frontier Cybersecurity Risks
Researcher, Frontier Cybersecurity Risks

OpenAI • Los Angeles (CA)

On-site
USD 295,000 - 445,000
Researcher, Frontier Cybersecurity Risks
Researcher, Frontier Cybersecurity Risks

Slope • San Francisco (CA)

On-site
USD 295,000 - 445,000
Researcher, Frontier Cybersecurity Risks
Researcher, Frontier Cybersecurity Risks

OpenAI • San Francisco (CA)

On-site
USD 295,000 - 445,000
Data Scientist, Safety
Data Scientist, Safety

OpenAI • New York (NY)

On-site
USD 230,000 - 325,000
Medical, dental, and vision insurance
401(k) retirement plan with employer match
Generous paid parental leave
+2
Fullstack Engineer, Safety Engineering
Fullstack Engineer, Safety Engineering

OpenAI • San Francisco (CA)

On-site
USD 210,000 - 325,000
Researcher, Frontier Cybersecurity Risks
Researcher, Frontier Cybersecurity Risks

CHEManager International • San Francisco (CA)

On-site
USD 170,000 - 210,000