AI Safety Red Teamer

🔥 0 minutes ago

🇺🇸 United States – Remote

💵 $70 - $84 / hour

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Mercor

Mercor

51 - 200 employees

Founded 2023

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

Mercor is a company for which no descriptive text was provided in the input. Additional information (products, services, target customers, or industry specifics) is needed to create an accurate summary and select appropriate industries.

📋 Description

• Design adversarial prompts to stress-test frontier AI models • Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports • Collaborate with AI researchers to improve model alignment, robustness, and safety • Work on projects focused on training and enhancing AI systems

🎯 Requirements

• Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field • Strong analytical reasoning, prompt design, and written communication skills • Experience designing adversarial prompts or evaluating frontier AI systems • Preferred: experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety • Preferred: familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies • Preferred: expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety • Must be able to work as an independent contractor • H1-B and STEM OPT candidates are not supported

🏖️ Benefits

• Fully remote role • Flexible own schedule • Weekly payments via Stripe or Wise • Projects may be extended, shortened, or concluded early depending on needs and performance • Competitive pay • Collaboration with leading AI researchers and safety teams • Opportunity to work on cutting-edge adversarial testing • Opportunity to influence AI safety and model development • Referral bonus of up to $340 per successful referral • Reasonable accommodations available upon request

Apply Now

Similar Jobs

🕒 2 days ago

Volga Partners

1001 - 5000

🤖 Artificial Intelligence

🤝 B2B

🏢 Enterprise

AI reviewer evaluating Italian-generated content and annotating language data. Providing linguistic feedback to improve artificial intelligence quality and reliability.

🗣️🇮🇹 Italian Required

🕒 2 days ago

10a Labs

11 - 50

🤖 Artificial Intelligence

🔒 Cybersecurity

☁️ SaaS

Biology SME reviewing biological content for 10a Labs’ AI safety and threat-intelligence platform. Applying scientific judgment to distinguish legitimate research from harmful or dual-use risks.

🕒 2 days ago

Globalme (now Summa Linguae Technologies)

51 - 200

🤖 Artificial Intelligence

🤝 B2B

💼 Consulting

Linguistic Analyst supporting AI language training and smart text improvements for a leading technology product. Evaluating datasets, analyzing linguistics, and validating quality with Product, Engineering, and Research teams.

🕒 3 days ago

Study.com

51 - 200

📚 Education

🛍️ eCommerce

AI tutorial creator producing narrated walkthroughs for Study.com’s generative-AI college courses. Demonstrating practical LLM workflows for non-technical learners.

🕒 3 days ago

Aston Carter

1001 - 5000

🎯 Recruiter

💼 Consulting

🤝 B2B

AI technical content specialist reviewing engineering and scientific outputs for Aston Carter’s AI training programs. Creating rigorous scenarios, solutions, and feedback to improve advanced AI systems.