Head of AI Safety in London

Head of AI Safety in London

London Full-Time 67000 - 80000 £ / year (est.) No working from home possible
M

At a Glance

  • Tasks: Lead AI safety initiatives, ensuring ethical practices and effective evaluation of AI systems.
  • Company: Moonshot, a mission-driven organisation focused on preventing online harms.
  • Benefits: 30 days annual leave, flexible holidays, private healthcare, and generous parental leave.
  • Other info: Join a collaborative team dedicated to tackling critical societal issues through innovative AI solutions.
  • Why this job: Make a real difference in AI safety while working with diverse teams and communities.
  • Qualifications: Experience in trust & safety or related fields, strong project management skills, and excellent communication.

The predicted salary is between 67000 - 80000 £ per year.

Description

Moonshot believes that marginalised people in society — including minority ethnic people, people from working class backgrounds, women, Disabled and LGBTQIA+ people — must be centred in the work we do.

We strongly encourage applications from people with these identities or who are members of other communities who are currently underrepresented in our workforce.

We know a diverse workforce will enable us to understand drivers behind violent extremism and online harms in an in-depth way and do better work to counter them.

About the role

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio.

The role combines Moonshot's expertise in violence prevention, behavioural risk, and online harms with the emerging practice of evaluating and improving the safety of AI systems.

The portfolio addresses harm categories including pathways to violence, extremism, child sexual exploitation, abuse and grooming (CSEA), mental health and crisis, and risks affecting children and teenagers.

The Head of AI Safety will serve as Moonshot's primary applied AI safety counterpart for frontier AI companies, governments, and regulators.

The role will work closely with model, policy, trust and safety, product, research, and engineering teams.

This is not an engineering or data-science role, but it is a hands-on position requiring the successful candidate to lead and participate directly in red teaming and adversarial evaluation, working in detail with evaluation methodologies, test scenarios, model responses, safety policies, and intervention frameworks.

The role holds responsibility for client and partner relationships, project and staff management, methodological quality, and business development.

The Head of AI Safety will build and maintain relationships across the wider AI safety ecosystem, including with governments, foundations, regulators, academics, researchers, and civil society organisations.

Your responsibilities will include

  • Applied AI Safety, Evaluation, and Advisory
  • Lead and quality-assure Moonshot's applied AI safety work across harm categories including pathways to violence, extremism, CSEA, abuse and grooming, mental health and crisis, and risks affecting children and teens, using methods such as red teaming and adversarial evaluation of AI systems.
  • Advise frontier AI companies on how to improve the safety of their models, products, policies, and intervention systems.
  • Translate insights from psychologists, child-safety specialists, violence-prevention practitioners, safeguarding experts, and other subject-matter experts into clear, actionable guidance for model safety, policy, product, research, and engineering teams.
  • Set the methodological approach for the portfolio, translating violence-prevention, safeguarding, and behavioural-risk expertise into structured and testable evaluation frameworks.
  • Lead and participate directly in red teaming and adversarial evaluation, working in detail with test scenarios, model responses, scoring criteria, safety policies, and evaluation results.
  • Identify patterns, edge cases, and potential safety failures, and develop practical recommendations for improving model behaviour and user protections.
  • Maintain rigour and clear documentation across the team's technical deliverables, suitable for technical, government, and foundation audiences.
  • Ensure work is delivered within a clear ethical framework and in compliance with contractual, legal, data protection, and ethics obligations.
  • Identify, manage, and escalate operational, reputational, delivery, and partnership risks.
  • Client & Partner Management
  • Serve as Moonshot's primary applied AI safety counterpart for frontier AI company partners, governments, regulators, and the wider ecosystem invested in AI safety.
  • Build trusted relationships with model, policy, trust and safety, product, research, and engineering teams.
  • Build and sustain relationships across the wider AI safety ecosystem, including governments, foundations, regulators, academics, researchers, civil society organisations, and specialist practitioners.
  • Represent Moonshot externally in meetings, briefings, workshops, and sector engagement, including with regulators and policymaker audiences.
  • Team Leadership & Management
  • Provide direct leadership, coaching, and management to Moonshot's AI safety team.
  • Foster a collaborative, accountable, and mission-driven team culture, with particular attention to wellbeing given the sensitive nature of the work.
  • Support workforce planning, performance management, and professional development across the team.
  • Ensure effective coordination with internal teams supporting the portfolio, including operations, finance, research, and technical teams.
  • Portfolio Development & Growth
  • Develop Moonshot's AI safety portfolio, identifying strategic opportunities, partnerships, and funding.
  • Lead proposal development, scoping, and renewals with technical credibility, using precise, defensible language suited to technical and government audiences.
  • Develop repeatable methodologies, service offerings, and partnerships that allow the portfolio to grow while maintaining methodological rigour and delivery quality.
  • Support external communications, publications, briefings, and thought leadership that establish Moonshot as a credible voice in applied AI safety.
  • Oversee project planning, staffing, budgeting, forecasting, and delivery timelines across the portfolio.

Requirements

Essential

  • Experience in trust & safety, online harms, or a closely related field such as violence prevention, safeguarding, or public health, and the ability to adapt that knowledge to AI systems.
  • Curiosity about AI and the ability to build technical fluency quickly, enough to engage credibly with technical counterparts at AI companies.

Much of this work is new, so comfort learning as you go matters more than existing AI safety expertise.

  • Experience designing research, evaluation frameworks, or interventions for harm categories such as violent extremism, CSEA, self-harm and crisis, or targeted violence.
  • Demonstrated experience managing projects, teams, budgets, partners, and clients, with strong people management skills.
  • Excellent written communication, with experience producing credible (not promotional) material for government, foundation, or enterprise audiences.
  • Comfort and demonstrated resilience working with highly sensitive or graphic content (CSEA, extremist material, crisis content), with awareness of wellbeing practices for this kind of work.
  • Strong judgment and the ability to navigate ambiguity, competing priorities, and sensitive stakeholder environments, including representing organisations externally.
  • Willingness to travel and work outside regular hours where needed to accommodate clients or respond to incidents.
  • Highly trustworthy, with discretion and diplomacy, and willing to undertake relevant security clearance procedures.
  • Experience supporting business development, grant funding, or procurement.
  • Commitment to Moonshot's mission.
  • We require and will check candidates' eligibility to work in the UK and pass any relevant security clearance procedures per the needs of clients.

Desirable

  • Direct experience in model safety, red teaming, or adversarial evaluation of LLMs or other AI systems.
  • Understanding of LLM architecture, safety tooling, or trust & safety policy.
  • Prior experience in child safety evaluation, teen-safety product work, or grooming and CSEA detection.
  • Familiarity with government or regulatory engagement, such as briefing officials or supporting policy submissions.
  • Experience with intervention or diversion programme design that can transfer to AI-mediated interventions.
  • Academic or applied background in radicalisation studies, forensic psychology, or violence risk assessment.
  • Familiarity with taxonomy or classifier development, including how testing data feeds a classifier.

Benefits

  • 30 days' paid annual leave, excluding bank holidays.
  • Flexible public holiday policy with the option to work bank holidays in exchange for a day off at another time.
  • Private healthcare package with access to specialist mental health cover, including coverage for partners and children.
  • Employee Assistance Programme providing access to mental health support.
  • Generous maternity and paternity leave: 26 weeks paid maternity leave, 8 weeks paid paternity leave.
  • All permanent employees are granted share options upon employment.

Head of AI Safety in London employer: Moonshot

Moonshot is an exceptional employer that fosters a dynamic work culture where innovation and collaboration thrive. As an OSINT Analyst, you will have the opportunity to engage in meaningful work that directly impacts the understanding of organised crime, while benefiting from professional growth opportunities and a supportive team environment. Located in the UK, we offer a unique chance to contribute to high-quality analytical reports and make a real difference in the field of online threat intelligence.

M

Contact Details:

Moonshot Recruitment Team

StudySmarter Expert Advice🤫

We think this is how you could land Head of AI Safety in London

Join Local Tech Meetups

Get out there and mingle with fellow developers by joining local tech meetups. It’s a fantastic way to meet people who might be working at Moonshot or know someone who does. Plus, you can pick up some trendy tech skills and trends while you're at it!

Contribute to Open Source Projects

Show off your coding chops by jumping into open-source projects. Not only does this give you practical experience, but it also gets you noticed in the dev community. You'll create a killer portfolio that speaks volumes about your skills to Moonshot.

Tap into Online Developer Communities

Don’t underestimate the power of online developer communities like GitHub, Stack Overflow, and even Reddit. Participate in discussions, share your projects, and build your visibility. We can often find opportunities through these channels that can lead to a full-time gig at companies like Moonshot.

Explore Job Boards Specifically for Tech Roles

Keep your eyes peeled on job boards that focus on tech roles. Sites like TechCareers or Stack Overflow Jobs can often have listings for companies like Moonshot that might not show up on broader job sites. Make it a habit to check these regularly, and don’t hesitate to apply directly through our website!

We think you need these skills to ace Head of AI Safety in London

Trust & Safety
Online Harms
Violence Prevention
Safeguarding
Public Health
AI Systems Knowledge
Research Design

Some tips for your application 🫡

Show off your coding skills:When applying for a software engineering role, it's super important to showcase your coding skills. Make sure your CV includes your tech stack, any relevant programming languages you’re comfortable with, and examples of projects you've worked on. If you have a GitHub profile, link it up! We love to see code in action.

Tailor your portfolio:For a full-time role, we’d expect to see some solid examples of your work in your portfolio. Make sure to include at least two or three projects that highlight your problem-solving skills and your ability to work with different technologies. Focus on the projects that are most relevant to the position at Moonshot.

Craft a killer cover letter:Your cover letter is your chance to stand out—make it personal! Explain why you want to work at Moonshot and how your skills align with the role. Show us your passion for software development. We dig enthusiastic candidates who understand the value of collaboration and continuous learning!

Be clear and concise:When it comes to writing your CV and cover letter, clarity is key. Avoid jargon that could confuse us and stick to simple, direct language. Highlight your achievements with quantifiable results where possible, and keep everything easy to read. A well-organised application goes a long way!

How to prepare for a job interview at Moonshot

Brush Up on Your Coding Skills

For a full-time software engineering role, it's crucial that we stay sharp with our coding abilities. Expect technical questions that might involve solving problems on the spot or discussing algorithms. Practise on platforms like LeetCode or HackerRank to get comfortable with the types of questions that often come up.

Know Your Tools and Frameworks

Make sure we’re well-acquainted with the tools and technologies listed in the job description. Familiarise ourselves with any specific frameworks or programming languages mentioned. If Moonshot uses React or Node.js, for instance, be ready to discuss how we’ve used them in previous projects or coursework.

Showcase Your Projects

Bring along a portfolio that highlights our best work. This could be code samples, GitHub repositories, or any side projects we’ve built. Make sure we can talk through our thought process for each project, especially the challenges we faced and how we solved them—this shows our problem-solving skills in action.

Prepare for Behavioural Questions

While technical skills are key, full-time positions also require cultural fit. Be ready to discuss our previous experiences and how we handle teamwork, conflict, and deadlines. Brush up on the STAR method—Situation, Task, Action, Result—to clearly articulate our past experiences when discussing how we've contributed to a team.