Head of Cyber Safety

Madrona Venture Labs

United States

On-site

USD 230,000 - 280,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

401k with up to 4% matching
28 days annual leave
Health, dental, and vision coverage
Catered lunches (Pittsburgh office)
Flexible work arrangements
Visa sponsorship available for rare/ep

Job summary

Gray Swan is seeking a seasoned cybersecurity leader in the United States to serve as the technical authority on AI-enabled cyber risk. You will design adversarial evaluations of frontier LLMs, translate cybersecurity expertise into scalable benchmarks, and drive safety standards across red-teaming, evaluation, and defenses.

You will mentor a world-class team and collaborate with AI labs and researchers to reduce catastrophic cyber risks from advanced AI systems.

Qualifications

  • Deep technical expertise in offensive cybersecurity, vulnerability research, exploit development, penetration testing, malware analysis, reverse engineering, or a closely related field.
  • Significant experience assessing advanced cyber threats, offensive tooling, or AI-enabled cyber capabilities, especially in critical infrastructure domains.
  • Hands-on experience conducting adversarial evaluations, AI red-teaming, LLM security research, or building evaluation datasets for frontier AI systems.
  • Comfortable operating at the intersection of cybersecurity research, AI safety, and machine learning engineering.
  • Thrive in highly ambiguous, fast-moving environments where you'll define strategy while building entirely new capabilities.
  • A builder who enjoys creating teams, infrastructure, and evaluation systems from scratch.

Responsibilities

  • Design and lead adversarial evaluations of frontier LLMs for offensive cyber capabilities, including vulnerability discovery, exploit development, malware generation, privilege escalation, social engineering, persistence, and autonomous cyber operations across text, agentic, and multimodal systems.
  • Partner closely with machine learning engineers to translate cybersecurity expertise into scalable benchmarks, classifiers, guardrails, automated detection systems, and evaluation infrastructure for both internal products and frontier AI lab deployments.
  • Develop and maintain Gray Swan’s catastrophic cyber harm taxonomy, continuously evolving cyber evaluation frameworks as frontier model capabilities rapidly advance.
  • Produce technical risk assessments and actionable recommendations for frontier AI labs, enterprise customers, and internal stakeholders, helping guide responsible model deployment and security mitigations.
  • Build, mentor, and lead a world-class team of cybersecurity subject matter experts while establishing scalable evaluation processes, quality standards, and technical infrastructure.
  • Represent Gray Swan as the company's cybersecurity authority, collaborating with frontier AI labs, security researchers, government partners, and the broader AI safety and cybersecurity communities.

Skills

Offensive cybersecurity
Vulnerability research
Adversarial AI security
AI red-teaming
LLM security research
ML/AI engineering collaboration
Python
Go
Rust

Tools

Python
Go
Rust

Job description

About Gray Swan

Gray Swan is on a mission to empower the world to use AI safely and securely. We evaluate AI models for the leading frontier labs along with building real-time threat detection and adaptive adversarial red teaming agents for teams deploying AI.

We're a team of approximately 50 people, well-funded, growing quickly. Our work directly influences how the world deploys AI agents and systems at scale..

Learn more about how we work.

The Role

Come build and lead Gray Swan's cyber safety capability from the ground up, serving as the technical authority on AI-enabled cyber risk across red-teaming, evaluation, benchmarking, defenses, and safety infrastructure development. You'll help define how frontier AI systems are evaluated for offensive cyber capabilities while partnering with leading AI labs to reduce real-world security risks.

This role sits at the intersection of offensive security, AI safety, and machine learning. You'll transform deep cybersecurity expertise into scalable evaluation methodologies, safety infrastructure, and automated defenses that help establish industry standards for frontier model security.

If you have deep expertise in offensive cybersecurity, vulnerability research, or adversarial AI security, experience evaluating frontier models, and is driven to reduce catastrophic cyber risks from increasingly capable AI systems, we'd love to hear from you.

What You’ll Do:
  • Design and lead adversarial evaluations of frontier LLMs for offensive cyber capabilities, including vulnerability discovery, exploit development, malware generation, privilege escalation, social engineering, persistence, and autonomous cyber operations across text, agentic, and multimodal systems.

  • Partner closely with machine learning engineers to translate cybersecurity expertise into scalable benchmarks, classifiers, guardrails, automated detection systems, and evaluation infrastructure for both internal products and frontier AI lab deployments.

  • Develop and maintain Gray Swan’s catastrophic cyber harm taxonomy, continuously evolving cyber evaluation frameworks as frontier model capabilities rapidly advance.

  • Produce technical risk assessments and actionable recommendations for frontier AI labs, enterprise customers, and internal stakeholders, helping guide responsible model deployment and security mitigations.

  • Build, mentor, and lead a world-class team of cybersecurity subject matter experts while establishing scalable evaluation processes, quality standards, and technical infrastructure.

  • Represent Gray Swan as the company's cybersecurity authority, collaborating with frontier AI labs, security researchers, government partners, and the broader AI safety and cybersecurity communities.

Who You Are:
  • Deep technical expertise in offensive cybersecurity, vulnerability research, exploit development, penetration testing, malware analysis, reverse engineering, or a closely related field through industry, research, or equivalent experience.

  • Significant experience assessing advanced cyber threats, offensive tooling, or AI-enabled cyber capabilities, especially in critical infrastructure domains.

  • Hands‑on experience conducting adversarial evaluations, AI red‑teaming, LLM security research, or building evaluation datasets for frontier AI systems.

  • Comfortable operating at the intersection of cybersecurity research, AI safety, and machine learning engineering.

  • Thrive in highly ambiguous, fast‑moving environments where you'll define strategy while building entirely new capabilities.

  • A builder who enjoys creating teams, infrastructure, and evaluation systems from scratch.

Bonus Points If You Have:
  • Experience developing machine learning models, AI security classifiers, or automated cyber detection systems.

  • Hands‑on experience red‑teaming frontier language models, jailbreaking, prompt injection research, or agentic AI evaluations.

  • Experience working with frontier AI labs, national security organizations, or leading cybersecurity research teams.

  • Background in threat intelligence, autonomous cyber operations, AI agent security, or AI governance.

  • Strong software engineering experience in Python, Go, Rust, or other systems programming languages.

If you don’t have 100% of these, you should still seriously consider applying. We care more about what you can do than your credentials.

You’ll Thrive Here If You:
  • Want visibility across the frontier of AI security by evaluating multiple frontier models and working directly with the world's leading AI labs.

  • Are excited to build the infrastructure that makes AI cyber safety scalable.

  • Feel motivated to reduce catastrophic cyber risks posed by increasingly capable AI systems.

  • Enjoy turning offensive security research into practical defenses that improve the safety of frontier AI.

  • Have a vision for what world‑class AI cyber safety should look like—and are excited to build it.

What We Offer:

We offer a competitive compensation package designed to reward impact and incentivize growth. Our compensation philosophy is informed by our current valuation and recent industry data.

Compensation: $230,000 - $280,000 plus performance based bonus and meaningful equity package

Benefits:

  • 401k with up to 4% matching

  • 28 days annual leave (vacation + holidays)

  • Health, dental, and vision coverage

  • Catered lunches (Pittsburgh office)

  • Flexible work arrangements

  • Visa sponsorship available for exceptional candidates

Interview Process

Application review. We read everything; we’ll respond within 10 days.

Online technical screen (15 min). Complete a simple, job-relevant exercise.

Intro call (30 min). We learn about you; you learn about us.

Technical interview (90 min). Live coding with some tasks requiring AIand others not.

Experience & culture interview (60 min). Conversational exploration of the skills fit.

Reference checks. We’ll reach out to 3-5 references that you provide.

Offer. If it’s mutual, we move fast.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Head of Biology
Head of Biology

Madrona Venture Labs • United States

On-site
USD 230,000 - 280,000
401k matching
Health, dental, vision
Paid vacation
+2
Software Engineer, Infrastructure
Software Engineer, Infrastructure

Madrona Venture Labs • Pittsburgh

On-site
USD 180,000 - 290,000
401k with matching
28 days annual leave
Health, dental, and vision
+3
Head of Biology
Head of Biology

Gray Swan AI • United States

On-site
USD 230,000 - 280,000
Health benefits
Equity package
Visa sponsorship for exceptional cand.
Head of Biology
Head of Biology

Gray Swan • United States

Hybrid
USD 230,000 - 280,000
401k matching
28 days leave
Health/dental/vision
+3
Software Engineer, Infrastructure
Software Engineer, Infrastructure

Gray Swan • United States

On-site
USD 180,000 - 290,000
401k with up to 4% matching
28 days annual leave
Health, dental, and vision coverage
+3
Strategic Account Executive – Frontier Labs
Strategic Account Executive – Frontier Labs

Gray Swan AI • United States

Hybrid
USD 157,000 - 192,000
401k with up to 4% matching
28 days annual leave
Health, dental, and vision coverage
+3
Software Engineer, Infrastructure
Software Engineer, Infrastructure

Gray Swan • Pittsburgh

Hybrid
USD 180,000 - 290,000
401k with up to 4% matching
28 days annual leave
Health, dental, and vision coverage
+3
Red Team Engineer
Red Team Engineer

Gray Swan AI • United States

On-site
USD 110,000 - 185,000
401k matching
Paid time off
Health insurance
+3
Machine Learning Engineer
Machine Learning Engineer

Gray Swan AI • Pittsburgh

On-site
USD 140,000 - 225,000
401k with up to 4% matching
28 days annual leave (vacation +bahold
Health, dental, and vision coverage
+3
Senior Software Engineer (Pittsburgh)
Senior Software Engineer (Pittsburgh)

Madrona Venture Labs • Pittsburgh

On-site
USD 180,000 - 220,000
401k matching
Paid vacation
Health insurance
+3