Search Remote Jobs

Senior Consultant, AI Safety

🔥 0 minutes ago

🌐 United Kingdom, Canada, +2 more countries – Remote

infoinfo

⏳ Contract/Temporary

🟠 Senior

🤖 Artificial Intelligence

🇬🇧 UK Skilled Worker Visa Sponsor

infoinfo

👻 Ghost score 16%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Moonshot

Moonshot

11 - 50 employees

🎖️ Defense

💼 Consulting

🔒 Cybersecurity

Defense • Consulting • Cybersecurity

Moonshot is a company that harnesses the power of the internet to address and resolve online harms globally. With a focus on both technology and human ethics, the company develops innovative methodologies and tools to identify and mitigate threats to communities, businesses, and governments, whether they are online or offline. Moonshot collaborates with experts in law enforcement, academia, cybersecurity, and government to provide actionable insights and interventions to prevent online violence and reinforce community safety. By deploying global campaigns and working with local support networks, Moonshot delivers measurable impacts and resilience against digital threats. Their services are extensive, encompassing threat monitoring, online violence prevention, and tech audits, among others, while ensuring data privacy and GDPR compliance.

📋 Description

• Conduct red teaming and adversarial evaluation of AI systems against defined harm categories. • Review model responses against harm and risk criteria and provide expert judgement. • Apply subject matter expertise to harm areas such as grooming and CSEA, radicalisation pathways, crisis signalling, or teen online safety. • Support the design of evaluation frameworks translating real-world harm knowledge into structured, testable criteria. • Contribute to intervention logic connecting at-risk users to appropriate support. • Draft methodology or findings for technical and government audiences. • Work closely with Moonshot's AI Safety team on the delivery of its AI Safety portfolio. • Deliver work within contractual, legal, data protection, and ethics obligations. • Role excludes client relationship management, team leadership, and business development responsibilities.

🎯 Requirements

• Experience in trust and safety, online harms, or a closely related field such as violence prevention, safeguarding, or public health, with the ability to apply that knowledge to AI systems. • Demonstrated experience designing research, evaluation frameworks, or interventions for harm categories such as violent extremism, CSEA, self-harm and crisis, or targeted violence. • Ability to translate real world knowledge of how a harm works into a way of testing whether an AI system handles it safely. • Comfort and demonstrated resilience working with highly sensitive or graphic content, including violence, extremist material, and crisis content, with awareness of wellbeing practices for this kind of work. • Strong written communication, able to produce credible, non promotional material for technical and government audiences. • Sound judgement working with ambiguity and sensitive material. • Availability for a close to full time commitment over approximately 6 weeks. • Willingness to undertake relevant security clearance procedures if required by the engagement. • Desirable: genuine depth in model safety, red teaming, or adversarial evaluation of LLMs or other AI systems. • Desirable: understanding of LLM architecture, safety tooling, or trust and safety policy. • Desirable: child safety evaluation, teen safety product work, or grooming and CSEA detection. • Desirable: intervention or diversion programme design transferable to AI-mediated interventions. • Desirable: government or regulatory engagement, such as briefing officials or supporting policy submissions. • Desirable: academic or applied background in radicalisation studies, forensic psychology, or violence risk assessment. • Desirable: taxonomy or classifier development, including how testing data feeds a classifier.

🏖️ Benefits

• Flexible working arrangements. • Opportunity to work on diverse, impactful projects. • Competitive consultancy rates. • Remote working options available.

Apply Now

Similar Jobs

🕒 2 days ago

Xenon Seven

11 - 50

💼 Consulting

🏥 Healthcare

🏭 Manufacturing

Agentforce AI Maestro architecting Salesforce agents, integrations, and LLM workflows for Xenon7’s European enterprise clients. Leading technical discovery, secure deployments, testing, and advisory engagements.

🕒 August 14

Aptura

1 - 1

🤖 Artificial Intelligence

🏥 Healthcare

💸 Finance

Physicians reviewing clinical outputs and designing evaluation tasks for frontier AI models. Improving model safety through specialty-based validation, grading criteria, and clinical feedback.

🕒 August 14

Welo Global

1001 - 5000

🤖 Artificial Intelligence

🤝 B2B

☁️ SaaS

Freelance Generative AI Analyst evaluating and annotating text, audio, images, and video for Welocalize’s global AI data company. Providing quality judgments and feedback under detailed project guidelines.

🕒 August 11

Cloudinary

201 - 500

☁️ SaaS

🔌 API

🛍️ eCommerce

GEO/AEO specialist growing Cloudinary’s visibility and citations across AI-powered search engines. Running weekly experiments, automating content workflows, and building authority through outreach.

🇬🇧 United Kingdom – Remote

💰 $100M Secondary Market on 2022-02

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

🕒 July 25

Staysure Group

501 - 1000

✈️ Travel

🏥 Healthcare

💼 Consulting

Owns lifecycle of conversational AI assistants across Group brand portfolio. Designing, building, and iterating on hybrid generative assistants in digital and voice channels.