AI Safety Expert, English, Finnish

Job not on LinkedIn

🔥 1 minute ago

🇺🇸 United States – Remote

💵 $48 - $62 / hour

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

👻 Ghost score 20%

infoinfo

🗣️🇫🇮 Finnish Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Weekday (YC W21)

Weekday (YC W21)

11 - 50 employees

Founded 2021

💼 Consulting

👥 HR Tech

☁️ SaaS

Consulting • HR Tech • SaaS

Weekday is a modern recruitment platform that combines AI technologies with a vast database of potential candidates, aiming to streamline the hiring process for companies in India. They offer various services, including a proactive outreach approach that helps employers connect with top talent, as well as tools for candidates to easily apply for jobs. Weekday's emphasis on candidate engagement through multiple channels, including email, WhatsApp, and phone calls, sets it apart in the competitive landscape of recruitment agencies.

📋 Description

• Red team conversational AI models and agents using jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation • Generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks • Follow taxonomies, benchmarks, and playbooks to keep testing consistent • Produce reproducible reports, datasets, and attack cases for customers • Uncover vulnerabilities automated tests miss • Deliver reproducible artifacts that strengthen customer AI systems • Expand evaluation coverage by testing more scenarios and reducing production surprises

🎯 Requirements

• Fluent/native fluency in English and Finnish • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing • Ability to probe AI systems adversarially, including jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation • Ability to generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks • Ability to follow taxonomies, benchmarks, and playbooks • Ability to produce reproducible reports, datasets, and attack cases • Ability to explain risks clearly to technical and non-technical stakeholders • Ability to adapt across projects and customers

🏖️ Benefits

• Participation in higher-sensitivity projects is optional • Clear guidelines for sensitive-content work • Wellness resources

Apply Now

Similar Jobs

🔥 12 hours ago

Mercor

51 - 200

AI Safety Red Teamer adversarially testing Mercor’s frontier AI systems. Identifying jailbreaks, hallucinations, and policy failures across high-risk domains.

🇺🇸 United States – Remote

💵 $70 - $84 / hour

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

🔥 12 hours ago

Mercor

51 - 200

AI Safety Red Teamer uncovering vulnerabilities in frontier AI models for Mercor. Designing adversarial tests and improving model safety with leading researchers.

🇺🇸 United States – Remote

💵 $70 - $84 / hour

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

🔥 12 hours ago

Aptura

1 - 1

🤖 Artificial Intelligence

🏥 Healthcare

💸 Finance

Radiologists annotating chest CT scans on a remote asynchronous platform. Supporting clinical AI models through precise structural markup of thoracic findings.

🕒 Yesterday

Mercor

51 - 200

AI Safety Red Teamer stress-testing frontier AI models for Mercor, which partners with AI labs to train and enhance systems. Identifying jailbreaks, hallucinations, and safety failures.

🇺🇸 United States – Remote

💵 $70 - $84 / hour

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

🕒 Yesterday

Mercor

51 - 200

AI safety red team experts probing Mercor-partnered frontier AI models for vulnerabilities. Generating reproducible attack data that improves model safety and reliability.

🇺🇸 United States – Remote

💵 $17 - $25 / hour

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

🗣️🇲🇾 Malay Required