Full-Time

Head of AI Safety

Updated on 8/18/2026

Moonshot (Online Safety)

Moonshot (Online Safety)

51-200 employees

Provides GDPR-compliant data management & privacy

Compensation Overview

CA$125k - CA$165k/yr

+ Share options

Toronto, ON, Canada

Remote

Remote within Ontario, Canada; willingness to travel and work outside regular hours may be required.

Category
Engineering Management (1)
Required Skills
LLM
Forecasting

Get referred to Moonshot (Online Safety)

See people who can refer or advise you

Requirements
  • Experience in trust and safety, online harms, violence prevention, safeguarding, public health, or a closely related field, with the ability to adapt that knowledge to AI systems.
  • Curiosity about AI and the ability to build technical fluency quickly enough to engage credibly with technical counterparts at AI companies.
  • Experience designing research, evaluation frameworks, or interventions for harm categories such as violent extremism, child sexual exploitation and abuse, self-harm and crisis, or targeted violence.
  • Demonstrated experience managing projects, teams, budgets, partners, and clients, with strong people-management skills.
  • Excellent written communication and experience producing credible, non-promotional material for government, foundation, or enterprise audiences.
  • Comfort and demonstrated resilience working with highly sensitive or graphic content, including child sexual exploitation and abuse, extremist material, and crisis content, with awareness of relevant wellbeing practices.
  • Strong judgment and the ability to navigate ambiguity, competing priorities, and sensitive stakeholder environments, including representing organizations externally.
  • Willingness to travel and work outside regular hours when needed to accommodate clients or respond to incidents.
  • Trustworthiness, discretion, and diplomacy, and willingness to undertake relevant security-clearance procedures.
  • Experience supporting business development, grant funding, or procurement.
  • Commitment to Moonshot's mission.
  • Eligibility to work in Canada and successful completion of a standard background check and any relevant client-required security-clearance procedures.
Responsibilities
  • Lead and quality-assure applied AI safety work across harm categories including pathways to violence, extremism, child sexual exploitation and abuse, abuse and grooming, mental health and crisis, and risks affecting children and teenagers, using red teaming and adversarial evaluation of AI systems.
  • Advise frontier AI companies on improving the safety of their models, products, policies, and intervention systems.
  • Translate insights from psychologists, child-safety specialists, violence-prevention practitioners, safeguarding experts, and other subject-matter experts into actionable guidance for model safety, policy, product, research, and engineering teams.
  • Set the methodological approach for the portfolio by translating violence-prevention, safeguarding, and behavioural-risk expertise into structured and testable evaluation frameworks.
  • Lead and participate directly in red teaming and adversarial evaluation, working with test scenarios, model responses, scoring criteria, safety policies, and evaluation results.
  • Identify patterns, edge cases, and potential safety failures, and develop recommendations for improving model behaviour and user protections.
  • Maintain rigorous documentation across technical deliverables for technical, government, and foundation audiences.
  • Ensure work is delivered within an ethical framework and complies with contractual, legal, data-protection, and ethics obligations.
  • Identify, manage, and escalate operational, reputational, delivery, and partnership risks.
  • Serve as the primary applied AI safety counterpart for frontier AI company partners, governments, regulators, and the wider AI safety ecosystem.
  • Build trusted relationships with model, policy, trust and safety, product, research, and engineering teams.
  • Build and sustain relationships with governments, foundations, regulators, academics, researchers, civil society organizations, and specialist practitioners.
  • Represent Moonshot externally in meetings, briefings, workshops, and sector engagement, including with regulators and policymakers.
  • Provide direct leadership, coaching, and management to the AI safety team.
  • Foster a collaborative, accountable, mission-driven team culture with attention to team wellbeing.
  • Support workforce planning, performance management, and professional development across the team.
  • Coordinate with internal operations, finance, research, and technical teams supporting the portfolio.
  • Develop the AI safety portfolio by identifying strategic opportunities, partnerships, and funding.
  • Lead proposal development, scoping, and renewals with technical credibility and precise, defensible language.
  • Develop repeatable methodologies, service offerings, and partnerships that support portfolio growth while maintaining methodological rigour and delivery quality.
  • Support external communications, publications, briefings, and thought leadership in applied AI safety.
  • Oversee project planning, staffing, budgeting, forecasting, and delivery timelines across the portfolio.
Desired Qualifications
  • Direct experience in model safety, red teaming, or adversarial evaluation of large language models or other AI systems.
  • Understanding of large language model architecture, safety tooling, or trust and safety policy.
  • Prior experience in child-safety evaluation, teen-safety product work, or grooming and child sexual exploitation and abuse detection.
  • Familiarity with government or regulatory engagement, such as briefing officials or supporting policy submissions.
  • Experience with intervention or diversion programme design transferable to AI-mediated interventions.
  • Academic or applied background in radicalization studies, forensic psychology, or violence risk assessment.
  • Familiarity with taxonomy or classifier development, including how testing data feeds a classifier.
Moonshot (Online Safety)

Moonshot (Online Safety)

View

Moonshot Team provides services that help organizations manage data and protect user privacy with a focus on GDPR compliance. Its offerings include tailored data management solutions, compliance audits, and ongoing support delivered through subscription plans and consulting fees. The product workflow involves assessing a client’s current data practices, implementing privacy controls, and continuous monitoring to maintain regulatory compliance over time. Moonshot differentiates itself by grounding its work in ethics, empathy, and human rights, following a Do No Harm philosophy to guide all operations. The company’s goal is to help clients meet regulatory requirements while safeguarding user privacy and data rights.

Company Size

51-200

Company Stage

Series A

Total Funding

$7M

Headquarters

London, United Kingdom

Founded

2015

Get referred to Moonshot (Online Safety)

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Chicago Sky and USOPC wins expand Moonshot into sports safety budgets.
  • March 2026 New Zealand programme opens recurring public-sector revenue.
  • Platform partnerships with Spotify deepen distribution and data access.

What critics are saying

  • Meta, X, and YouTube can change moderation APIs, weakening Moonshot's delivery overnight.
  • Sports contracts concentrate revenue; Chicago Sky and USOPC procurement cycles remain fragile.
  • AI safety vendors commoditize detection, crushing margins by 2027 and collapsing differentiation.

What makes Moonshot (Online Safety) unique

  • Moonshot pairs national-security tooling with athlete and platform safety programs.
  • Spotify partnership and USOPC Team USA Safe Online validate specialized abuse-response workflows.
  • New Zealand's first nationwide online violence prevention programme shows government-grade deployment.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Health Insurance

Mental Health Support

Parental Leave

Stock Options

Remote Work Options

Flexible Work Hours

Paid Vacation

401(k) Company Match

Dental Insurance

Vision Insurance

Life Insurance

Disability Insurance

Growth & Insights and Company News

Headcount

6 month growth

-2%

1 year growth

0%

2 year growth

1%
eNews Park Forest
Jul 1st, 2025
WNBA's Chicago Sky Announces First-of-Its-Kind Partnership with Moonshot to Protect Players from Online Threats and Abuse

CHICAGO, July 1, 2025 /PRNewswire/ - Today, the WNBA's Chicago Sky announced an innovative partnership with Moonshot, a global leader in countering online threats, to leverage national security technology to keep the team's full roster safe.

CoinCu
Jun 1st, 2025
Moonshot Launches LOUD Token with $11.6 Million Value

Moonshot has launched the Solana-based LOUD token on June 1, 2025, with a current market value of $11.6 million.

CryptoTimes
Jan 18th, 2025
Trump Memecoin Pushes Moonshot to Top 10 on U.S. App Store

The excitement that followed the launch helped Moonshot, a rival to Pump.fun, which is run on the Solana blockchain.

PR Newswire
Sep 26th, 2024
New Initiative Empowers Young People To Help Curb Hate-Fueled Violence By Their Peers

FBI report detailing an increase in hate crimes underscores the need for the new UP End Hate campaign. PITTSBURGH, Sept. 26, 2024 /PRNewswire/ -- The Eradicate Hate Global Summit, the most comprehensive international anti-hate gathering in the world, is proud to announce the launch of an initiative aimed at curbing the troubling escalation of young people involved in acts of hate-fueled violence, including mass shootings. "The new UP End Hate campaign is the nation's first comprehensive initiative aimed at giving young people, ages 12-22, the skills and resources to prevent acts of violence by their peers," said Brette Steele, President of the Eradicate Hate Global Summit. "Young people are often the first to see or hear warning signs from some of their peers who may be considering turning to violence, so we want to empower them to recognize those signals and then take action before their peers make that dangerous turn."

Google
Feb 16th, 2023
Google returns to the Munich Security Conference

The initiative proved effective – so effective in fact, that this week Headwayservices announced Headwayservices is expanding it to Germany, in partnership with Moonshot and local experts.