Head of AI Safety

Job not on LinkedIn

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Moonshot

Moonshot

11 - 50 employees

🎖️ Defense

💼 Consulting

🔒 Cybersecurity

Defense • Consulting • Cybersecurity

Moonshot is a company that harnesses the power of the internet to address and resolve online harms globally. With a focus on both technology and human ethics, the company develops innovative methodologies and tools to identify and mitigate threats to communities, businesses, and governments, whether they are online or offline. Moonshot collaborates with experts in law enforcement, academia, cybersecurity, and government to provide actionable insights and interventions to prevent online violence and reinforce community safety. By deploying global campaigns and working with local support networks, Moonshot delivers measurable impacts and resilience against digital threats. Their services are extensive, encompassing threat monitoring, online violence prevention, and tech audits, among others, while ensuring data privacy and GDPR compliance.

📋 Description

• Lead and quality-assure Moonshot's applied AI safety work across violence, extremism, CSEA, abuse and grooming, mental health and crisis, and child and teen risk categories. • Advise frontier AI companies on improving model, product, policy, and intervention safety. • Translate subject-matter expertise into actionable guidance for model safety, policy, product, research, and engineering teams. • Set methodological approaches and develop structured, testable evaluation frameworks. • Lead and participate directly in red teaming and adversarial evaluation of AI systems. • Identify safety failures and develop recommendations for model behaviour and user protections. • Maintain rigorous documentation and ensure compliance with legal, data protection, contractual, and ethical obligations. • Manage operational, reputational, delivery, and partnership risks. • Serve as Moonshot's primary applied AI safety counterpart for partners, governments, regulators, and the wider ecosystem. • Build relationships with AI company teams, governments, foundations, academics, researchers, civil society organizations, and practitioners. • Represent Moonshot externally in meetings, briefings, workshops, and sector engagement. • Lead, coach, and manage the AI safety team. • Support workforce planning, performance management, professional development, and team wellbeing. • Coordinate with operations, finance, research, and technical teams. • Develop the AI safety portfolio through strategic opportunities, partnerships, and funding. • Lead proposal development, scoping, and renewals. • Develop repeatable methodologies, service offerings, and partnerships. • Support communications, publications, briefings, and thought leadership. • Oversee project planning, staffing, budgeting, forecasting, and delivery timelines.

🎯 Requirements

• Experience in trust & safety, online harms, violence prevention, safeguarding, or public health, with ability to adapt knowledge to AI systems. • Curiosity about AI and ability to build technical fluency quickly. • Experience designing research, evaluation frameworks, or interventions for violent extremism, CSEA, self-harm and crisis, or targeted violence. • Experience managing projects, teams, budgets, partners, and clients, with strong people management skills. • Excellent written communication for government, foundation, or enterprise audiences. • Resilience working with highly sensitive or graphic content, with awareness of wellbeing practices. • Strong judgment navigating ambiguity, competing priorities, and sensitive stakeholder environments. • Willingness to travel and work outside regular hours when needed. • Trustworthiness, discretion, diplomacy, and willingness to undertake security clearance procedures. • Experience supporting business development, grant funding, or procurement. • Commitment to Moonshot's mission. • Eligibility to work in Canada. • Required to pass a standard background check and relevant security clearance procedures per client needs. • Desirable: direct experience in model safety, red teaming, or adversarial evaluation of LLMs or other AI systems. • Desirable: understanding of LLM architecture, safety tooling, or trust & safety policy. • Desirable: child safety evaluation, teen-safety product work, or grooming and CSEA detection experience. • Desirable: government or regulatory engagement experience. • Desirable: intervention or diversion programme design experience. • Desirable: academic or applied background in radicalization studies, forensic psychology, or violence risk assessment. • Desirable: familiarity with taxonomy or classifier development and testing data.

🏖️ Benefits

• 25 days paid vacation leave, plus Statutory Holiday • Flexible public holiday policy with the option to work statutory holidays in exchange for a day off at another time. • Group healthcare package, including coverage for partners and children (80% Co-Insurance). • HSA is restricted to mental health practitioners only • Dental & Vision Insurance (80% Co-Insurance). • Life & LTD Disability Insurance. • 24/7 access to counselling via our Employee Assistance Program. • Generous maternity and paternity leave: 26 weeks paid maternity leave, 8 weeks paid paternity leave. • All permanent employees are granted share options upon employment.

Apply Now

Similar Jobs

🔥 14 hours ago

TELUS Digital

201 - 500

💼 Consulting

📣 Marketing

📦 Logistics

CXAI transformation director shaping customer-experience innovation at TELUS Digital, a digital product consultancy and TELUS division. Leading transformation roadmaps, ROI cases, communications, and AI capability-building across a 75,000-person operation.

🕒 July 10

NICE

5001 - 10000

☁️ SaaS

🤖 Artificial Intelligence

📡 Telecommunications

Principal Business Consultant focusing on AI solutions and digital transformations at NICE. Leading consulting projects and ensuring business readiness for optimal adoption of AI solutions.

🕒 June 15

Toptal

1001 - 5000

💼 Consulting

📣 Marketing

🎯 Recruiter

Vice President of AI Services leading AI-powered ITSM solutions for clients. Overseeing practice strategy, client engagement, and team development in a fully remote setup.

🕒 June 10

Instacart

1001 - 5000

🍽️ Food & Beverage

📦 Logistics

🛍️ eCommerce

AI Engagement Manager responsible for orchestrating AI engagements with B2B partners. Overseeing delivery precision and managing partner relationships within Instacart's Enterprise Solutions team.

🇨🇦 Canada – Remote

💵 $185k - $195k / year

💰 $232M Venture Round on 2021-11

⏰ Full Time

🟠 Senior

🔴 Lead

🤖 Artificial Intelligence

🕒 May 29

JDPA LIMITED

-

🚘 Automotive

🛡️ Insurance

💼 Consulting

Lead AI Enablement at JD Power, driving adoption of AI tools across the enterprise. Chair AI Steering Committee and build internal tools for effective AI integration.