Search Remote Jobs

Head of AI Safety

Job not on LinkedIn

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Moonshot

Moonshot

11 - 50 employees

🎖️ Defense

💼 Consulting

🔒 Cybersecurity

Defense • Consulting • Cybersecurity

Moonshot is a company that harnesses the power of the internet to address and resolve online harms globally. With a focus on both technology and human ethics, the company develops innovative methodologies and tools to identify and mitigate threats to communities, businesses, and governments, whether they are online or offline. Moonshot collaborates with experts in law enforcement, academia, cybersecurity, and government to provide actionable insights and interventions to prevent online violence and reinforce community safety. By deploying global campaigns and working with local support networks, Moonshot delivers measurable impacts and resilience against digital threats. Their services are extensive, encompassing threat monitoring, online violence prevention, and tech audits, among others, while ensuring data privacy and GDPR compliance.

📋 Description

• Lead and quality-assure Moonshot's applied AI safety work across violence, extremism, CSEA, abuse and grooming, mental health and crisis, and child and teen risk categories. • Advise frontier AI companies on improving model, product, policy, and intervention safety. • Translate specialist insights into actionable guidance for model safety, policy, product, research, and engineering teams. • Set methodological approaches and structured, testable evaluation frameworks. • Lead and participate in red teaming and adversarial evaluation using test scenarios, model responses, scoring criteria, safety policies, and evaluation results. • Identify safety failures and develop recommendations for model behavior and user protections. • Maintain rigorous documentation for technical, government, and foundation audiences. • Ensure ethical, legal, contractual, data protection, and compliance obligations are met. • Manage operational, reputational, delivery, and partnership risks. • Serve as Moonshot's primary applied AI safety counterpart for AI companies, governments, regulators, and ecosystem partners. • Build relationships with technical teams, governments, foundations, academics, researchers, civil society organizations, and practitioners. • Represent Moonshot in meetings, briefings, workshops, and sector engagement. • Lead, coach, and manage the AI safety team; support workforce planning, performance management, and professional development. • Coordinate with operations, finance, research, and technical teams. • Develop the AI safety portfolio, strategic opportunities, partnerships, and funding. • Lead proposal development, scoping, renewals, communications, publications, briefings, and thought leadership. • Develop repeatable methodologies and service offerings while maintaining rigor and delivery quality. • Oversee project planning, staffing, budgeting, forecasting, and delivery timelines.

🎯 Requirements

• Experience in trust & safety, online harms, violence prevention, safeguarding, public health, or a closely related field. • Ability to adapt relevant harm-prevention knowledge to AI systems. • Curiosity about AI and ability to build technical fluency quickly. • Experience designing research, evaluation frameworks, or interventions for violent extremism, CSEA, self-harm and crisis, or targeted violence. • Experience managing projects, teams, budgets, partners, and clients. • Strong people management skills. • Excellent written communication and experience producing credible material for government, foundation, or enterprise audiences. • Resilience and comfort working with highly sensitive or graphic content. • Strong judgment and ability to navigate ambiguity, competing priorities, and sensitive stakeholder environments. • Willingness to travel and work outside regular hours when needed. • Trustworthiness, discretion, diplomacy, and willingness to undertake relevant security clearance procedures. • Experience supporting business development, grant funding, or procurement. • Commitment to Moonshot's mission. • Eligibility to work in the US. • Ability to pass relevant security clearance procedures per client needs. • Desirable: experience in model safety, red teaming, or adversarial evaluation of LLMs or other AI systems. • Desirable: understanding of LLM architecture, safety tooling, or trust & safety policy. • Desirable: child safety evaluation, teen-safety product work, or grooming and CSEA detection experience. • Desirable: government or regulatory engagement experience. • Desirable: intervention or diversion program design experience. • Desirable: academic or applied background in radicalization studies, forensic psychology, or violence risk assessment. • Desirable: familiarity with taxonomy or classifier development and testing data.

🏖️ Benefits

• 15 days paid vacation leave, plus Federal holidays and 1 day additional paid leave for Native American Heritage Day. • Flexible public holiday policy with the option to work federal holidays in exchange for a day off at another time. • Full private healthcare package, including coverage for partners and children. • Dental & Vision Insurance. • Life & Disability Insurance. • 24/7 access to free counseling via our Employee Assistance Program. • 3% matched 401k contributions. • 401(k) Roth Contributions. • Generous maternity and paternity leave: 26 weeks paid maternity leave, 8 weeks paid paternity leave. • All permanent employees are granted share options upon employment.

Apply Now

Similar Jobs

🔥 5 hours ago

Teach For All

51 - 200

📚 Education

🤝 Non-profit

🌍 Social Impact

Chief Learning Officer advancing AI-driven educational transformation across Teach For All’s global network. Building AI partnerships, learning agendas, and responsible innovation across 60+ countries.

🔥 9 hours ago

Harbor

501 - 1000

⚖️ Legal

📣 Marketing

💼 Consulting

Director leading production AI engagements for Harbor, a legal-industry strategy, technology, operations, and intelligence services provider. Shaping client strategy, operating models, integrations, and delivery.

🔥 10 hours ago

Vultr

201 - 500

🤖 Artificial Intelligence

🤝 B2B

🔧 Hardware

AI Cluster Architect designing power-aware GPU clusters for Vultr’s global cloud infrastructure platform. Optimizing networking, thermal capacity, and GPU density across large-scale deployments.

🇺🇸 United States – Remote

💵 $165k - $185k / year

💰 $329M Debt Financing - Vultr on 2025-06

⏰ Full Time

🟠 Senior

🔴 Lead

🤖 Artificial Intelligence

🔥 11 hours ago

Entegris

5001 - 10000

Entegris Director leading AI and digital process innovation for semiconductor-materials R&D. Defining intelligent workflows, digital twins, governance, and scalable scientific solutions.

🔥 14 hours ago

TELUS Digital

201 - 500

💼 Consulting

📣 Marketing

📦 Logistics

TELUS Digital, a global digital product consultancy, seeks a transformation director for CXAI programs and organizational change. Connecting AI investments to contact-center metrics, roadmaps, ROI, capability building, and executive outcomes.