Head of AI Safety
- Location
- Greater London, England, United Kingdom
Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioural risk, and online harms with the emerging practice of evaluating and improving the safety of AI systems. The portfolio addresses harm categories including pathways to violence … actionable guidance for model safety, policy, product, research, and engineering teams. Set the methodological approach for the portfolio, translating violence‐prevention, safeguarding, and behavioural‐risk expertise into structured and testable evaluation frameworks. Lead and participate directly in red teaming and adversarial evaluation, working in detail with test scenarios, model ...