Head of AI Safety

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Moonshot

Moonshot

11 - 50 employees

🎖️ Defense

💼 Consulting

🔒 Cybersecurity

Defense • Consulting • Cybersecurity

Moonshot is a company that harnesses the power of the internet to address and resolve online harms globally. With a focus on both technology and human ethics, the company develops innovative methodologies and tools to identify and mitigate threats to communities, businesses, and governments, whether they are online or offline. Moonshot collaborates with experts in law enforcement, academia, cybersecurity, and government to provide actionable insights and interventions to prevent online violence and reinforce community safety. By deploying global campaigns and working with local support networks, Moonshot delivers measurable impacts and resilience against digital threats. Their services are extensive, encompassing threat monitoring, online violence prevention, and tech audits, among others, while ensuring data privacy and GDPR compliance.

📋 Description

• Lead the delivery, development, and growth of Moonshot's AI Safety portfolio • Lead and quality-assure applied AI safety work across violence, extremism, CSEA, abuse and grooming, mental health and crisis, and child and teen risk categories • Advise frontier AI companies on improving model, product, policy, and intervention safety • Translate subject-matter expertise into actionable guidance for model safety, policy, product, research, and engineering teams • Set methodological approaches and develop structured, testable evaluation frameworks • Lead and participate directly in red teaming and adversarial evaluation • Identify safety failures and develop recommendations for improving model behaviour and user protections • Maintain rigorous documentation and ensure ethical, legal, contractual, data protection, and compliance requirements • Manage operational, reputational, delivery, and partnership risks • Serve as Moonshot's primary applied AI safety counterpart for AI companies, governments, regulators, and ecosystem partners • Build relationships with technical teams, governments, foundations, regulators, academics, researchers, civil society organisations, and practitioners • Represent Moonshot externally in meetings, briefings, workshops, and sector engagement • Lead, coach, and manage the AI safety team • Support workforce planning, performance management, professional development, and team wellbeing • Coordinate with operations, finance, research, and technical teams • Develop the AI safety portfolio through strategic opportunities, partnerships, and funding • Lead proposal development, scoping, renewals, and business development • Develop repeatable methodologies, service offerings, and partnerships • Support communications, publications, briefings, and thought leadership • Oversee project planning, staffing, budgeting, forecasting, and delivery timelines

🎯 Requirements

• Experience in trust & safety, online harms, violence prevention, safeguarding, or public health, with ability to adapt knowledge to AI systems • Curiosity about AI and ability to build technical fluency quickly • Experience designing research, evaluation frameworks, or interventions for violent extremism, CSEA, self-harm and crisis, or targeted violence • Demonstrated experience managing projects, teams, budgets, partners, and clients • Strong people management skills • Excellent written communication for government, foundation, or enterprise audiences • Resilience and comfort working with highly sensitive or graphic content • Awareness of wellbeing practices for sensitive-content work • Strong judgment and ability to navigate ambiguity, competing priorities, and sensitive stakeholder environments • Willingness to travel and work outside regular hours when needed • Trustworthiness, discretion, diplomacy, and willingness to undertake relevant security clearance procedures • Experience supporting business development, grant funding, or procurement • Commitment to Moonshot's mission • Eligibility to work in the UK • Ability to pass relevant security clearance procedures • Direct experience in model safety, red teaming, or adversarial evaluation of LLMs or other AI systems (desirable) • Understanding of LLM architecture, safety tooling, or trust & safety policy (desirable) • Child safety evaluation, teen-safety product work, or grooming and CSEA detection experience (desirable) • Government or regulatory engagement experience (desirable) • Intervention or diversion programme design experience (desirable) • Academic or applied background in radicalisation studies, forensic psychology, or violence risk assessment (desirable) • Familiarity with taxonomy or classifier development and testing data workflows (desirable)

🏖️ Benefits

• 30 days' paid annual leave, excluding public holidays • Flexible public holiday policy with the option to work public holidays in exchange for a day off at another time • Private healthcare package with access to specialist mental health cover, including coverage for partners and children • Dental and Vision Insurance • Life Insurance & Income Protection • Employee Assistance Programme providing access to mental health support • 26 weeks paid maternity leave • 8 weeks paid paternity leave • Share options granted to all permanent employees upon employment

Apply Now

Similar Jobs

🕒 July 27

Holafly

501 - 1000

✈️ Travel

📡 Telecommunications

👥 B2C

Leading AI and automation strategy for Holafly's customer experience. Championing traveler journeys and mentoring a high-impact team to ensure seamless connectivity.

🇮🇪 Ireland – Remote

💰 $112.9k Seed Round - Holafly on 2019-06

⏰ Full Time

🔴 Lead

🤖 Artificial Intelligence