AI Safety Expert, English – Bengali

Job not on LinkedIn

🔥 0 minutes ago

🇺🇸 United States – Remote

💵 $20 - $22 / hour

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Mercor

Mercor

51 - 200 employees

Founded 2023

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

Mercor is a company for which no descriptive text was provided in the input. Additional information (products, services, target customers, or industry specifics) is needed to create an accurate summary and select appropriate industries.

📋 Description

• Red-team conversational AI models and agents • Test jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation • Annotate AI failures, classify vulnerabilities, and flag systemic risks • Apply taxonomies, benchmarks, and playbooks to maintain consistent testing • Produce reproducible reports, datasets, and attack cases for customers • Identify vulnerabilities that automated tests miss • Expand evaluation coverage and reduce production surprises • Help Mercor customers strengthen the safety, robustness, and trustworthiness of their AI systems • Work on projects training and enhancing frontier AI systems

🎯 Requirements

• Fluent/native fluency in English and Bengali required • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing • Ability to probe AI systems adversarially, including jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation • Ability to generate human data by annotating failures, classifying vulnerabilities, and flagging systemic risks • Ability to follow taxonomies, benchmarks, and playbooks • Ability to produce reproducible reports, datasets, and attack cases • Ability to explain risks clearly to technical and non-technical stakeholders • Adaptability across projects and customers • H1-B and STEM OPT candidates are not supported • Independent contractor engagement

🏖️ Benefits

• Fully remote role that can be completed on your own schedule • Payments are weekly on Stripe or Wise based on services rendered • Higher-sensitivity projects are optional • Clear guidelines and wellness resources for higher-sensitivity projects • Competitive pay • Referral bonus of up to $90 for each successful referral • No limit on referrals (restrictions may apply)

Apply Now

Similar Jobs

🔥 3 hours ago

Welo Global

1001 - 5000

🤖 Artificial Intelligence

🤝 B2B

☁️ SaaS

Generative AI Analyst evaluating and annotating AI-generated content for Welo Data, a global AI data company. Reviewing multilingual data quality and providing structured feedback.

🗣️🇨🇳 Chinese Required

🔥 3 hours ago

Gramian Consulting

2 - 10

💼 Consulting

📦 Logistics

📣 Marketing

Film and television expert evaluating AI-generated responses for Gramian Consultancy, an IT professional-services and engineering talent-solutions consultancy. Creating prompts, benchmarks, and evidence-based factuality reviews.

🔥 3 hours ago

Gramian Consulting

2 - 10

💼 Consulting

📦 Logistics

📣 Marketing

Video Games Domain Reviewer evaluating gaming prompts and tasks for Gramian Consultancy’s AI model quality projects. Checking factual accuracy, reasoning, completeness, and guideline compliance.

🕒 Yesterday

The Browser Company

11 - 50

💼 Consulting

📣 Marketing

AI Prototyper exploring frontier AI for Dia, The Browser Company’s next-generation browser. Building and evaluating prototypes that guide product strategy and engineering handoff.

🇺🇸 United States – Remote

💵 $100 - $150 / hour

💰 $13M Series A on 2021-05

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

🕒 Yesterday

Terac

1 - 10

🤖 Artificial Intelligence

🤝 B2B

Terac, which recruits vetted human experts for AI research, runs a paid study on designing challenging prompts. Participants test ChatGPT failures and create grading rubrics.