Head of AI Safety & Responsible Tech

Moonshot

Greater London

On-site

GBP 67,000 - 80,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

30 days' annual leave
Flexible public holiday policy
Private healthcare
Employee Assistance Programme
Maternity and paternity leave
Share options

Job summary

Moonshot in the United Kingdom is seeking a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio, combining violence prevention, behavioural risk, and online harms with evaluating AI system safety.

The role acts as Moonshot's primary safety counterpart for frontier AI companies, governments, and regulators, working with model, policy, trust and safety, product, research, and engineering teams to advance practical, ethical safeguards.

Qualifications

  • Experience in trust & safety, online harms, or related field and ability to adapt to AI systems.
  • Curiosity about AI with ability to engage credibly with technical teams.
  • Experience designing research, evaluation frameworks, or interventions for harm categories.
  • Experience managing projects, teams, budgets, partners, and clients.
  • Excellent written communication for government, foundation, or enterprise audiences.
  • Comfort handling highly sensitive or graphic content with wellbeing practices.
  • Strong judgment navigating ambiguity and competing priorities.
  • Willingness to travel and work outside regular hours when needed.
  • Discretion and diplomacy, with readiness for security procedures.
  • Experience supporting business development, grant funding, or procurement.
  • Commitment to Moonshot's mission.

Responsibilities

  • Lead and quality‑assure applied AI safety work across harm categories using red teaming and adversarial evaluation.
  • Advise frontier AI companies on improving model safety, policies, and interventions.
  • Translate insights from safety experts into actionable guidance for product, research, and engineering teams.
  • Set the methodological approach for the portfolio and develop structured evaluation frameworks.
  • Lead red teaming and adversarial evaluation with test scenarios and scoring criteria.
  • Identify patterns and edge cases, and provide practical recommendations for model behaviour and user protections.
  • Maintain rigorous documentation for technical, government, and foundation audiences.
  • Ensure work complies with contractual, legal, data protection, and ethics obligations.
  • Identify, manage, and escalate operational and partnership risks.

Skills

Trust & safety
AI safety curiosity
Research design
Project management
Written communication
Risk assessment
Ambiguity tolerance
Stakeholder management
Security clearance
Business development support
Mission alignment

Job description

Moonshot in the United Kingdom is seeking a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio, combining violence prevention, behavioural risk, and online harms with evaluating AI system safety.

The role acts as Moonshot's primary safety counterpart for frontier AI companies, governments, and regulators, working with model, policy, trust and safety, product, research, and engineering teams to advance practical, ethical safeguards.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Head of AI Safety & Ethical AI Leadership
Head of AI Safety & Ethical AI Leadership

Moonshotteam • Greater London

Hybrid
GBP 67,000 - 80,000
30 days paid annual leave
Flexible public holiday policy
Private healthcare package
+3
Head of AI Safety
Head of AI Safety

Moonshot • Greater London

On-site
GBP 67,000 - 80,000
30 days' annual leave
Flexible public holiday policy
Private healthcare
+3
Head of AI Safety
Head of AI Safety

Moonshotteam • Greater London

On-site
GBP 67,000 - 80,000
30 days paid annual leave
Flexible public holiday policy
Private healthcare package
+3
Senior AI Safety Red Team Consultant (Remote)
Senior AI Safety Red Team Consultant (Remote)

Moonshot • Greater London

Remote
GBP 148,000 - 221,000
Flexible working arrangements
Diverse, impactful projects
Competitive consultancy rates
+1
Chief AI Safety & Responsible Tech Leader
Chief AI Safety & Responsible Tech Leader

Faculty • Greater London

Hybrid
GBP 100,000 - 140,000
Unlimited Annual Leave Policy
Private healthcare and dental
Enhanced parental leave
+3
Senior Freelance Consultant, AI Safety
Senior Freelance Consultant, AI Safety

Moonshot • Greater London

Remote
GBP 148,000 - 221,000
Flexible working arrangements
Diverse, impactful projects
Competitive consultancy rates
+1
Staff Software Engineer, AI Safety & Safeguards
Staff Software Engineer, AI Safety & Safeguards

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 160,000
Comprehensive health insurance
Fertility benefits via Carrot Fertilty
Paid parental leave 22 weeks
+1
AI Safety Lead — Research, Evaluation & Governance
AI Safety Lead — Research, Evaluation & Governance

Best AI Tools Wiki • Greater London

Hybrid
GBP 180,000 - 240,000
Relocation to London supported
Private healthcare for family
Generous research budget
+2
Head of AI Safety & Strategic R&D
Head of AI Safety & Strategic R&D

Faculty AI • Greater London

Hybrid
GBP 180,000 - 240,000
Unlimited Annual Leave Policy
Private healthcare and dental
Enhanced parental leave
+2
AI Safety Director — Strategic R&D Leader
AI Safety Director — Strategic R&D Leader

Faculty • Greater London

Hybrid
GBP 80,000 - 120,000
Unlimited Annual Leave Policy
Private healthcare and dental
Enhanced parental leave
+2