
51 - 200 employees
Founded 2023
🔥 Funding within the last year
đź’° $350M Series C - Mercor on 2025-10
Mercor is a company for which no descriptive text was provided in the input. Additional information (products, services, target customers, or industry specifics) is needed to create an accurate summary and select appropriate industries.
🔥 0 minutes ago
🇺🇸 United States – Remote
đź’µ $60 - $70 / hour
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
Founded 2023
🔥 Funding within the last year
đź’° $350M Series C - Mercor on 2025-10
Mercor is a company for which no descriptive text was provided in the input. Additional information (products, services, target customers, or industry specifics) is needed to create an accurate summary and select appropriate industries.
• Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality • Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains • Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking • Identify unsafe outputs, hallucinations, reasoning failures, and policy violations • Provide structured feedback to improve model alignment and safety performance • Collaborate with AI researchers and safety teams on ongoing evaluation initiatives • Work on projects training and enhancing frontier AI systems
• Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline • 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field • Excellent written English • Critical thinking and analytical reasoning skills • Ability to consistently evaluate nuanced and policy-sensitive scenarios • Experience with AI Safety, RLHF, SFT, Trust & Safety, or AI evaluation preferred • Familiarity with safety policies, content moderation, or evaluation rubric development preferred • Experience reviewing complex, high-risk, or ambiguous content preferred • Must not require H1-B or STEM OPT sponsorship
• Fully remote role • Flexible own schedule • Projects can be extended, shortened, or concluded early depending on needs and performance • Weekly payments via Stripe or Wise based on services rendered • Competitive pay • Collaboration with leading AI researchers, engineers, and safety teams • Opportunity to shape frontier AI models used by millions worldwide • Up to $400 for each successful referral, with no limit on referrals • Reasonable accommodations upon request
Apply Now🔥 1 hour ago
AI Change Lead driving strategy, communications and adoption for NextLink’s global pharmaceutical AI transformation programme. Coordinating governance, learning pathways, champions and stakeholder engagement.
đź•’ 2 days ago
French-speaking Mac freelancer testing an enterprise client’s AI audio application. Evaluating generated audio, application behavior, bugs, and instruction adherence.
🇺🇸 United States – Remote
đź’° $6.2M Series A - Lifted on 2021-06
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
🗣️🇫🇷 French Required
đź•’ 2 days ago
German-speaking Mac freelancer evaluating an enterprise client’s AI voice application. Recording outputs, testing behavior, and identifying bugs across up to 10 short tasks.
🇺🇸 United States – Remote
đź’° $6.2M Series A - Lifted on 2021-06
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
🗣️🇩🇪 German Required
đź•’ 2 days ago
AI safety specialist red teaming conversational AI for 24-MAG’s remote consulting platform. Developing jailbreaks, classifying vulnerabilities, and producing structured evaluation data and reports.
🇺🇸 United States – Remote
đź’µ $45 - $60 / hour
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
🗣️🇸🇪 Swedish Required
đź•’ 2 days ago
AI Safety Specialist red-teaming conversational AI models for 24-MAG’s remote consulting platform. Identifying vulnerabilities, annotating failures, and producing structured safety evaluation data.
🇺🇸 United States – Remote
đź’µ $45 - $60 / hour
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
🗣️🇳🇱 Dutch Required