×
Register Here to Apply for Jobs or Post Jobs. X
More jobs:

Head of AI Safety

Job in Greater London, London, Greater London, W1B, England, UK
Listing for: Moonshotteam
Full Time position
Listed on 2026-09-13
Job specializations:
  • Business
    AI Evaluation
Salary/Wage Range or Industry Benchmark: 67000 - 80000 GBP Yearly GBP 67000.00 80000.00 YEAR
Job Description & How to Apply Below
Location: Greater London

Moonshot believes that marginalised people in society — including minority ethnic people, people from working class backgrounds, women, Disabled and LGBTQIA+ people — must be centred in the work we do. We strongly encourage applications from people with these identities or who are members of other communities who are currently underrepresented in our workforce. We know a diverse workforce will enable us to understand drivers behind violent extremism and online harms in an in-depth way and do better work to counter them.

About the role

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioural risk, and online harms with the emerging practice of evaluating and improving the safety of AI systems. The portfolio addresses harm categories including pathways to violence, extremism, child sexual exploitation, abuse and grooming (CSEA), mental health and crisis, and risks affecting children and teenagers.

The Head of AI Safety will serve as Moonshot's primary applied AI safety counterpart for frontier AI companies, governments, and regulators. The role will work closely with model, policy, trust and safety, product, research, and engineering teams. This is not an engineering or data-science role, but it is a hands‑on position requiring the successful candidate to lead and participate directly in red teaming and adversarial evaluation, working in detail with evaluation methodologies, test scenarios, model responses, safety policies, and intervention frameworks.

The role holds responsibility for client and partner relationships, project and staff management, methodological quality, and business development. The Head of AI Safety will build and maintain relationships across the wider AI safety ecosystem, including with governments, foundations, regulators, academics, researchers, and civil society organisations.

Your responsibilities will include:

Applied AI Safety, Evaluation, and Advisory

  • Lead and quality‑assure Moonshot's applied AI safety work across harm categories including pathways to violence, extremism, CSEA, abuse and grooming, mental health and crisis, and risks affecting children and teens, using methods such as red teaming and adversarial evaluation of AI systems.
  • Advise frontier AI companies on how to improve the safety of their models, products, policies, and intervention systems.
  • Translate insights from psychologists, child‑safety specialists, violence‑prevention practitioners, safeguarding experts, and other subject‑matter experts into clear, actionable guidance for model safety, policy, product, research, and engineering teams.
  • Set the methodological approach for the portfolio, translating violence‑prevention, safeguarding, and behavioural‑risk expertise into structured and testable evaluation frameworks.
  • Lead and participate directly in red teaming and adversarial evaluation, working in detail with test scenarios, model responses, scoring criteria, safety policies, and evaluation results.
  • Identify patterns, edge cases, and potential safety failures, and develop practical recommendations for improving model behaviour and user protections.
  • Maintain rigour and clear documentation across the team's technical deliverables, suitable for technical, government, and foundation audiences.
  • Ensure work is delivered within a clear ethical framework and in compliance with contractual, legal, data protection, and ethics obligations.
  • Identify, manage, and escalation operational, reputational, delivery, and partnership risks.

Client & Partner Management

  • Serve as Moonshot's primary applied AI safety counterpart for frontier AI company partners, governments, regulators, and…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary