
Head of AI Safety
Posted Aug 7

Posted Aug 7
This is a fully remote position, open to applicants in Canada.
• Oversee and ensure the quality of Moonshot's applied AI safety initiatives across categories such as violence, extremism, CSEA, abuse and grooming, mental health and crisis, as well as child and teen risk.
• Provide guidance to frontier AI companies on enhancing the safety of their models, products, policies, and interventions.
• Convert specialized knowledge into practical recommendations for model safety, policy, product development, research, and engineering teams.
• Establish methodological strategies and create structured, testable evaluation frameworks.
• Directly engage in red teaming and adversarial evaluations of AI systems.
• Detect safety shortcomings and formulate recommendations for model behavior and user protection.
• Uphold meticulous documentation and ensure adherence to legal, data protection, contractual, and ethical standards.
• Manage operational, reputational, delivery, and partnership risks.
• Act as Moonshot's key applied AI safety liaison for partners, governments, regulators, and the broader ecosystem.
• Cultivate relationships with AI company teams, government bodies, foundations, academics, researchers, civil society organizations, and practitioners.
• Represent Moonshot in external meetings, briefings, workshops, and sector engagements.
• Lead, mentor, and oversee the AI safety team.
• Assist in workforce planning, performance management, professional development, and team wellbeing.
• Collaborate with operations, finance, research, and technical teams.
• Expand the AI safety portfolio through strategic initiatives, partnerships, and funding opportunities.
• Spearhead proposal development, scoping, and renewal processes.
• Create repeatable methodologies, service offerings, and strategic partnerships.
• Support communications, publications, briefings, and thought leadership efforts.
• Supervise project planning, staffing, budgeting, forecasting, and delivery schedules.
• Proven experience in trust & safety, online harms, violence prevention, safeguarding, or public health, with the ability to adapt this knowledge to AI systems.
• A keen interest in AI and the capability to quickly develop technical fluency.
• Experience in designing research, evaluation frameworks, or interventions targeting violent extremism, CSEA, self-harm and crisis, or targeted violence.
• Proven track record in managing projects, teams, budgets, partners, and clients, along with strong people management abilities.
• Exceptional written communication skills tailored for government, foundation, or enterprise audiences.
• Resilience in handling highly sensitive or graphic content, with an understanding of wellbeing practices.
• Strong judgment in navigating ambiguity, competing priorities, and sensitive stakeholder environments.
• Willingness to travel and work outside of regular hours as required.
• Trustworthy, discreet, diplomatic, and prepared to undergo security clearance procedures.
• Experience in supporting business development, grant funding, or procurement processes.
• Commitment to advancing Moonshot's mission.
• Eligibility to work in Canada.
• Must pass a standard background check and relevant security clearance processes as per client requirements.
• Desirable: direct experience in model safety, red teaming, or adversarial evaluation of LLMs or other AI systems.
• Desirable: a grasp of LLM architecture, safety tools, or trust & safety policies.
• Desirable: experience in evaluating child safety, teen-safety product development, or detecting grooming and CSEA.
• Desirable: experience in engagement with government or regulatory bodies.
• Desirable: background in designing intervention or diversion programs.
• Desirable: academic or practical experience in radicalization studies, forensic psychology, or violence risk assessment.
• Desirable: familiarity with taxonomy or the development and testing of classifiers and data.
• 25 days of paid vacation leave, in addition to Statutory Holidays.
• Flexible public holiday policy allowing the option to work statutory holidays in exchange for a day off at a later date.
• Comprehensive group healthcare package covering partners and children (80% Co-Insurance).
• HSA restricted to mental health practitioners only.
• Dental & Vision Insurance (80% Co-Insurance).
• Life & Long-Term Disability Insurance.
• 24/7 access to counseling through our Employee Assistance Program.
• Generous parental leave policies: 26 weeks of paid maternity leave and 8 weeks of paid paternity leave.
• Permanent employees are granted share options upon hiring.
Progressive Leasing
apna
apna
Texas Research International
Get handpicked remote jobs straight to your inbox weekly.