Senior GenAI Safety Researcher

Job not on LinkedIn

🔥 0 minutes ago

🗽 New York – Remote

infoinfo

💵 $105k - $115k / year

⏰ Full Time

🟠 Senior

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of ActiveFence

ActiveFence

201 - 500 employees

🔒 Cybersecurity

🤖 Artificial Intelligence

💰 $400M Series B on 2021-07

Cybersecurity • Artificial Intelligence

ActiveFence is a Trust and Safety provider for online platforms, protecting platforms and their users from malicious behavior and content. Trust and Safety teams of all sizes rely on ActiveFence to keep their users safe from the widest spectrum of online harms, unwanted content, and malicious behavior, including child safety and exploitation, disinformation, hate speech, terror, nudity, fraud, and more. We offer a full stack of capabilities with our deep intelligence research, AI-driven harmful content detection, and online content moderation platform. Protecting over three billion users globally everyday in over 100 languages, ActiveFence lets people interact and thrive online.

📋 Description

• Architect rigorous, scalable testing methodologies and red-teaming frameworks for foundational models, multimodal systems, and AI agents • Develop prompt strategies across risk domains including hate speech, misinformation, IP and copyright infringement, and child safety • Research emerging jailbreak tactics, prompt injection techniques, and circumvention strategies • Provide technical and policy insight to support the Program Lead in project scoping, risk assessment, and strategy • Document, synthesize, and expand Alice’s internal AI Safety knowledge base • Standardize best practices, taxonomies, and research findings across the team • Mentor junior analysts and elevate analytical standards • Own engagement lifecycles from planning and methodology design through execution, QA, and final delivery • Oversee complex, multi-language datasets across multiple abuse areas • Partner with engineering, product, policy, and client-facing teams to communicate findings and inform mitigation strategies

🎯 Requirements

• 5+ years of experience in AI Safety, Responsible AI, Trust & Safety, or aligned research domains • Proven expertise in research design and building qualitative or quantitative evaluation methodologies for GenAI • Strong domain expertise in content risks, including toxicity, copyright, misinformation, and safety policy violations • Track record of project ownership, leading deliverables end-to-end with high attention to detail in fast-paced, variable environments • Deep familiarity with modern Generative AI architectures, prompt engineering, red-teaming, and AI agents • Strong communication skills to act as a core subject-matter contact for program leads, internal teams, and clients • Published research in academia, industry whitepapers, or a research institute (nice-to-have) • Hands-on experience evaluating multimodal systems such as Text-to-Image, Text-to-Video, and Audio (nice-to-have) • Experience mentoring, leading, or QAing junior analysts and researchers (nice-to-have)

Apply Now

Similar Jobs

🔥 6 hours ago

DecisionPoint Corporation

51 - 200

🎖️ Defense

💼 Consulting

📦 Logistics

Emerging technology researcher evaluating AI, cloud, cybersecurity, and automation innovations for DecisionPoint’s federal and DoD mission environments. Supporting vendor assessments, pilots, and technology adoption decisions.

🕒 Yesterday

GreyNoise Intelligence

11 - 50

🔒 Cybersecurity

Threat Researcher hunting malicious cyber activity for GreyNoise Intelligence, a startup providing real-time Internet threat intelligence. Producing actionable intelligence to help customers detect and disrupt adversaries.

🕒 August 14

Dhaka Technologies Limited Company

2 - 10

🤝 B2B

🏢 Enterprise

🔒 Cybersecurity

Senior BI Data Researcher advancing AI, machine learning, and distributed data systems for an education and research client. Building scalable tools and architectures across diverse data types.

🇺🇸 United States – Remote

💵 $140k - $160k / year

⏰ Full Time

🟠 Senior

🕒 August 14

ImagineX

201 - 500

💼 Consulting

🏥 Healthcare

📦 Logistics

Lead Researcher shaping AI and enterprise product decisions for ImagineX’s client engagements. Running mixed-method research, service design, and executive-level facilitation.

🇺🇸 United States – Remote

💰 Private equity on 2023-11

⏰ Full Time

🟠 Senior

🕒 August 11

American Institutes for Research

1001 - 5000

🏥 Healthcare

💼 Consulting

📚 Education

Healthcare quality measurement researcher leading CMS-focused policy research at AIR. Developing measures, analyzing complex datasets, and guiding expert panels, reports, QA/QC, budgets, and proposals.