Chief AI Safety & Adversarial Evaluation Lead

Moonshotteam

Washington, Denver, Atlanta (District of Columbia, CO, GA)

Hybrid

USD 110,000 - 145,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Healthcare package
Dental & Vision Insurance
Life & Disability Insurance
401k with matching

Job summary

Moonshot is seeking a Head of AI Safety to lead delivery, development, and growth of our AI Safety portfolio. You will work with model, policy, trust and safety, product, research, and engineering teams to drive applied safety initiatives, red teaming, and advisory services for frontier AI companies and regulators.

You will build relationships with governments, foundations, and industry partners while managing projects, budgets, and staff.

Qualifications

  • Experience in trust & safety or violence prevention and applying to AI systems.
  • Ability to learn quickly and engage credibly with technical teams.
  • Experience designing research or evaluation frameworks for harm categories such as violence or online safety.

Responsibilities

  • Lead and quality‑assure applied AI safety work across harm categories using red teaming and adversarial evaluation.
  • Advise frontier AI companies on improving safety of models, products, policies, and interventions.
  • Translate insights from safety and safeguarding experts into actionable guidance for technical and policy teams.
  • Set methodologies translating expertise into structured evaluation frameworks.
  • Lead red teaming and evaluation with test scenarios, model responses, and scoring criteria.
  • Identify safety edge cases and provide practical recommendations to improve model behavior.
  • Maintain rigorous documentation for technical, government, and foundation audiences.
  • Ensure ethical and legal compliance in all work; manage risks and partnerships.

Skills

Applied AI Safety
Red Teaming
Adversarial Evaluation
Stakeholder Management
Policy & Compliance
Government Engagement

Job description

Moonshot is seeking a Head of AI Safety to lead delivery, development, and growth of our AI Safety portfolio. You will work with model, policy, trust and safety, product, research, and engineering teams to drive applied safety initiatives, red teaming, and advisory services for frontier AI companies and regulators.

You will build relationships with governments, foundations, and industry partners while managing projects, budgets, and staff.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Head of AI Safety & Adversarial Evaluation
Head of AI Safety & Adversarial Evaluation

Moonshot • Denver (CO)

On-site
USD 110,000 - 120,000
15 days paid vacation
Federal holidays
Healthcare package
+6
Head of AI Safety & Risk Governance with Equity Options
Head of AI Safety & Risk Governance with Equity Options

Moonshot • Washington

On-site
USD 110,000 - 145,000
Private healthcare package
Dental & Vision Insurance
Life & Disability Insurance
+2
Head of AI Safety & Risk - Equity Options
Head of AI Safety & Risk - Equity Options

Moonshot • Atlanta (GA)

On-site
USD 110,000 - 120,000
15 days vacation leave
Federal holidays and additional leave
Private healthcare package
+6
Director, AI Safety & Risk — Equity Options
Director, AI Safety & Risk — Equity Options

Moonshot • Washington

On-site
USD 110,000 - 120,000
15 days paid vacation
Private healthcare
Dental & Vision Insurance
+4
Head of AI Safety
Head of AI Safety

Moonshot • Denver (CO)

On-site
USD 110,000 - 120,000
15 days paid vacation
Federal holidays
Healthcare package
+6
Head of AI Safety
Head of AI Safety

Moonshot • Washington

On-site
USD 110,000 - 120,000
15 days paid vacation
Private healthcare
Dental & Vision Insurance
+4
Head of AI Safety
Head of AI Safety

Moonshotteam • Washington, Denver (CO), Atlanta (GA)

On-site
USD 110,000 - 145,000
Healthcare package
Dental & Vision Insurance
Life & Disability Insurance
+1
Head of AI Safety
Head of AI Safety

Moonshot • Washington

On-site
USD 110,000 - 145,000
Private healthcare package
Dental & Vision Insurance
Life & Disability Insurance
+2
Head of AI Safety
Head of AI Safety

Moonshot • Atlanta (GA)

On-site
USD 110,000 - 120,000
15 days vacation leave
Federal holidays and additional leave
Private healthcare package
+6
AI Governance & Policy Research Lead
AI Governance & Policy Research Lead

Safetytalent • San Francisco (CA)

On-site
USD 100,000 - 130,000