Head of AI Safety leading applied AI system evaluations and advisory work at Moonshot.
Managing teams, partnerships, and portfolio growth across violence prevention and online harms.
Responsibilities
Lead and quality-assure Moonshot's applied AI safety work across violence, extremism, CSEA, abuse and grooming, mental health and crisis, and child and teen risk categories.
Advise frontier AI companies on improving model, product, policy, and intervention safety.
Translate subject-matter expertise into actionable guidance for model safety, policy, product, research, and engineering teams.
Set methodological approaches and develop structured, testable evaluation frameworks.
Lead and participate directly in red teaming and adversarial evaluation of AI systems.
Identify safety failures and develop recommendations for model behaviour and user protections.
Maintain rigorous documentation and ensure compliance with legal, data protection, contractual, and ethical obligations.
Manage operational, reputational, delivery, and partnership risks.
Serve as Moonshot's primary applied AI safety counterpart for partners, governments, regulators, and the wider ecosystem.
Build relationships with AI company teams, governments, foundations, academics, researchers, civil society organizations, and practitioners.
Represent Moonshot externally in meetings, briefings, workshops, and sector engagement.
Lead, coach, and manage the AI safety team.
Support workforce planning, performance management, professional development, and team wellbeing.
Coordinate with operations, finance, research, and technical teams.
Develop the AI safety portfolio through strategic opportunities, partnerships, and funding.
Lead proposal development, scoping, and renewals.
Develop repeatable methodologies, service offerings, and partnerships.
Support communications, publications, briefings, and thought leadership.
Oversee project planning, staffing, budgeting, forecasting, and delivery timelines.
Requirements
Experience in trust & safety, online harms, violence prevention, safeguarding, or public health, with ability to adapt knowledge to AI systems.
Curiosity about AI and ability to build technical fluency quickly.
Experience designing research, evaluation frameworks, or interventions for violent extremism, CSEA, self-harm and crisis, or targeted violence.
Experience managing projects, teams, budgets, partners, and clients, with strong people management skills.
Excellent written communication for government, foundation, or enterprise audiences.
Resilience working with highly sensitive or graphic content, with awareness of wellbeing practices.
Strong judgment navigating ambiguity, competing priorities, and sensitive stakeholder environments.
Willingness to travel and work outside regular hours when needed.
Trustworthiness, discretion, diplomacy, and willingness to undertake security clearance procedures.
Experience supporting business development, grant funding, or procurement.
Commitment to Moonshot's mission.
Eligibility to work in Canada.
Required to pass a standard background check and relevant security clearance procedures per client needs.
Desirable: direct experience in model safety, red teaming, or adversarial evaluation of LLMs or other AI systems.
Desirable: understanding of LLM architecture, safety tooling, or trust & safety policy.
Desirable: child safety evaluation, teen-safety product work, or grooming and CSEA detection experience.
Desirable: government or regulatory engagement experience.
Desirable: intervention or diversion programme design experience.
Desirable: academic or applied background in radicalization studies, forensic psychology, or violence risk assessment.
Desirable: familiarity with taxonomy or classifier development and testing data.
Benefits
25 days paid vacation leave, plus Statutory Holiday
Flexible public holiday policy with the option to work statutory holidays in exchange for a day off at another time.
Group healthcare package, including coverage for partners and children (80% Co-Insurance).
HSA is restricted to mental health practitioners only
Dental & Vision Insurance (80% Co-Insurance).
Life & LTD Disability Insurance.
24/7 access to counselling via our Employee Assistance Program.
Search +AI Discovery Manager leading enterprise SEO, AI - search strategy, and client advisory for dentsu, a global integrated marketing agency. Mentoring North American SEO teams and shaping innovative search methodologies.
Director leading Homebase’s data engineering, data science, and applied AI teams. Building trusted data infrastructure and AI - powered tools for 150,000+ small businesses.
AI Solution Specialist helping greenhouse growers adopt Source.ag’s AI - powered crop technology. Building dashboards, automations, and custom tools while translating greenhouse insights into product improvements.
Senior Advisor designing and integrating generative AI solutions for Desjardins, North America’s largest cooperative financial group. Driving adoption through training, change management and scalable AI practices.
Strategy and management consultant creating complex business scenarios and evaluating AI outputs. Supporting Gramian Consultancy’s IT talent solutions through rigorous consulting analysis and executive deliverables.
Gartner Director Analyst shaping AI - driven technology operations guidance for CIOs. Conducting research, advising executives, and presenting modernization strategies to global clients.
Romanian AI Tutor training SpaceXAI’s Grok for multilingual speech and voice interactions. Annotating audio, translating content, and evaluating accents, pronunciation, and prosody.
AI Tutor training SpaceXAI’s Grok with Slovenian multilingual audio data. Annotating speech, recording voices, and localizing text to improve global voice interactions.
AI Agent Implementation Specialist building AI agents and administering Marketo for Extreme Networks’ cloud - driven networking solutions. Supporting MarTech integrations, campaign workflows, troubleshooting, and marketing automation across US time zones.
Global Public Policy Manager shaping compute, infrastructure, export - control, and sovereign AI policy. Cohere builds security - first enterprise AI models and products.