Moonshot is seeking a Head of AI Safety to lead the development, delivery, and expansion of its AI Safety portfolio. This role integrates Moonshot's expertise in violence prevention, behavioral risk, and online harms with the emerging field of AI system safety evaluation and enhancement. The portfolio addresses critical harm categories such as pathways to violence, extremism, child sexual exploitation, abuse and grooming (CSEA), mental health crises, and risks impacting children and teenagers. The Head of AI Safety will act as Moonshot's primary applied AI safety contact for leading AI companies, governments, and regulatory bodies. The position requires close collaboration with model, policy, trust and safety, product, research, and engineering teams. While not an engineering or data science role, it is a hands-on position involving direct participation in red teaming and adversarial evaluations, deep engagement with evaluation methodologies, test scenarios, model responses, safety policies, and intervention frameworks. Key responsibilities include managing client and partner relationships, staff and project oversight, ensuring methodological quality, and business development. The role also involves building and maintaining relationships within the broader AI safety ecosystem, including governmental bodies, foundations, regulators, academics, researchers, and civil society organizations. This is a remote position, but candidates must be based in Ontario, Canada, to comply with employment and regulatory requirements.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed