AI Safety Red Teamer

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Mercor

Mercor

51 - 200 employees

Founded 2023

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

Mercor is a company for which no descriptive text was provided in the input. Additional information (products, services, target customers, or industry specifics) is needed to create an accurate summary and select appropriate industries.

📋 Description

• Design adversarial prompts to stress-test frontier AI models • Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports • Collaborate with AI researchers to improve model alignment, robustness, and safety • Work on projects focused on training and enhancing AI systems

🎯 Requirements

• Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field • Strong analytical reasoning, prompt design, and written communication skills • Experience designing adversarial prompts or evaluating frontier AI systems • Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety preferred • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies preferred • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety preferred • Must be unable to require H1-B or STEM OPT support at this time

🏖️ Benefits

• Fully remote work • Flexible own schedule • Weekly payments via Stripe or Wise based on services rendered • Opportunity to work alongside leading AI researchers and safety teams • Cutting-edge adversarial testing projects • Influence how AI systems respond to real-world safety challenges • Competitive payment • Referral opportunity earning up to $340 per successful referral • Reasonable accommodations upon request

Apply Now

Similar Jobs

🔥 11 hours ago

Crunchbase

51 - 200

☁️ SaaS

🤝 B2B

Independent consultant building AI-enabled products, prototypes, and workflow automations for Crunchbase, a private-company intelligence platform. Delivering defined solutions during an approximately two-month remote engagement.

🔥 14 hours ago

Terac

1 - 10

🤖 Artificial Intelligence

🤝 B2B

Remote contract contributors sourcing legally usable PDFs in Telugu, Odia, Gujarati, Malayalam, Japanese, or Korean. Supporting Terac’s AI text recognition and generation research.

🗣️🇯🇵 Japanese Required

🗣️🇰🇷 Korean Required

🕒 Yesterday

24-MAG

2 - 10

🤝 B2B

💼 Consulting

PhD chemistry consultant evaluating AI-generated scientific responses for accuracy, safety, and responsible handling of sensitive content. Remote, part-time project work for 24-MAG LLC.

🕒 Yesterday

Study.com

51 - 200

📚 Education

🛍️ eCommerce

Generative AI lesson creator producing narrated screen-capture tutorials for Study.com, an online education platform. Demonstrating practical LLM workflows for college-level business and responsible-AI courses.

🕒 2 days ago

NextLink Group

201 - 500

💼 Consulting

🏥 Healthcare

📦 Logistics

AI Change Lead driving strategy, communications and adoption for NextLink’s global pharmaceutical AI transformation programme. Coordinating governance, learning pathways, champions and stakeholder engagement.