AI Safety Expert – English, Danish

Job not on LinkedIn

🔥 12 hours ago

🇺🇸 United States – Remote

💵 $48 - $62 / hour

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

👻 Ghost score 0%

infoinfo

🗣️🇩🇰 Danish Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Mercor

Mercor

51 - 200 employees

Founded 2023

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

Mercor is a company for which no descriptive text was provided in the input. Additional information (products, services, target customers, or industry specifics) is needed to create an accurate summary and select appropriate industries.

📋 Description

• Red-team conversational AI models and agents through jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation • Generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks • Apply taxonomies, benchmarks, and playbooks to maintain consistent testing • Produce reproducible reports, datasets, and attack cases for customers • Uncover vulnerabilities that automated tests miss • Expand evaluation coverage and reduce production surprises • Help Mercor customers strengthen the safety, robustness, and trustworthiness of their AI systems

🎯 Requirements

• Fluent/native fluency in English and Danish is required • Prior red teaming experience, including AI adversarial work, cybersecurity, or socio-technical probing • Ability to probe AI systems adversarially and push them to breaking points • Experience using frameworks, taxonomies, benchmarks, or playbooks for structured testing • Ability to explain risks clearly to technical and non-technical stakeholders • Adaptability across projects and customers • Nice-to-have specialties: adversarial ML, jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction, penetration testing, exploit development, reverse engineering, harassment/disinformation probing, abuse analysis, conversational AI testing, psychology, acting, or unconventional adversarial writing • Must be eligible to work without H1-B or STEM OPT support

🏖️ Benefits

• Fully remote role that can be completed on your own schedule • Payments weekly on Stripe or Wise based on services rendered • Higher-sensitivity projects are optional and supported by clear guidelines and wellness resources • Competitive pay • Opportunity to collaborate with leading researchers • Referral payments of up to $250 for each successful referral

Apply Now

Similar Jobs

🔥 15 hours ago

Volga Partners

1001 - 5000

🤖 Artificial Intelligence

🤝 B2B

🏢 Enterprise

AI reviewer evaluating Italian-generated content and annotating language data. Providing linguistic feedback to improve artificial intelligence quality and reliability.

🗣️🇮🇹 Italian Required

🔥 19 hours ago

10a Labs

11 - 50

🤖 Artificial Intelligence

🔒 Cybersecurity

☁️ SaaS

Biology SME reviewing biological content for 10a Labs’ AI safety and threat-intelligence platform. Applying scientific judgment to distinguish legitimate research from harmful or dual-use risks.

🔥 20 hours ago

Globalme (now Summa Linguae Technologies)

51 - 200

🤖 Artificial Intelligence

🤝 B2B

💼 Consulting

Linguistic Analyst supporting AI language training and smart text improvements for a leading technology product. Evaluating datasets, analyzing linguistics, and validating quality with Product, Engineering, and Research teams.

🕒 Yesterday

Study.com

51 - 200

📚 Education

🛍️ eCommerce

AI tutorial creator producing narrated walkthroughs for Study.com’s generative-AI college courses. Demonstrating practical LLM workflows for non-technical learners.

🕒 2 days ago

Aston Carter

1001 - 5000

🎯 Recruiter

💼 Consulting

🤝 B2B

AI technical content specialist reviewing engineering and scientific outputs for Aston Carter’s AI training programs. Creating rigorous scenarios, solutions, and feedback to improve advanced AI systems.