
501 - 1000 employees
Founded 2014
đź’Ľ Consulting
🏥 Healthcare
📦 Logistics
Consulting • Healthcare • Logistics
Handshake is a platform that connects students with potential employers and career centers, offering tools to streamline the recruiting process for its users. It provides features such as job postings, talent engagement suites, and event management for both virtual and in-person recruiting events. Handshake is widely used by Fortune 100 companies and offers integrations with existing automated tracking systems. It aims to build early career networks and supports a diverse range of Gen Z talent with a focus on improving hiring outcomes through proprietary data and advanced recruiting strategies.
🔥 9 hours ago
🇺🇸 United States – Remote
đź’µ $65 - $158 / hour
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
🦅 H1B Visa Sponsor
đź‘» Ghost score 0%
Improve your chances of getting an interview by checking your resume score before you apply.

501 - 1000 employees
Founded 2014
đź’Ľ Consulting
🏥 Healthcare
📦 Logistics
Consulting • Healthcare • Logistics
Handshake is a platform that connects students with potential employers and career centers, offering tools to streamline the recruiting process for its users. It provides features such as job postings, talent engagement suites, and event management for both virtual and in-person recruiting events. Handshake is widely used by Fortune 100 companies and offers integrations with existing automated tracking systems. It aims to build early career networks and supports a diverse range of Gen Z talent with a focus on improving hiring outcomes through proprietary data and advanced recruiting strategies.
• Evaluate whether AI models appropriately handle queries related to chemical, biological, radiological, nuclear, and explosive threats • Design technically grounded adversarial prompts testing whether models provide meaningful uplift toward CBRNE threats • Evaluate model outputs for technical accuracy and dangerous information • Probe dual-use knowledge boundaries involving scientific, medical, industrial, and weapons applications • Test multi-step and multi-turn attack chains • Score model responses against structured harm taxonomies and severity rubrics • Document findings with clear technical reasoning • Distinguish open-literature information from genuine operational uplift • Contribute to CBRNE-specific evaluation frameworks and threat models • Collaborate with red teamers, AI researchers, and policy teams to translate findings into model improvements • Stay current on model capabilities, jailbreak techniques, and relevant domain developments
• Graduate-level education or equivalent professional experience in a relevant CBRNE field (chemistry, biochemistry, microbiology, virology, nuclear physics, radiochemistry, materials science, munitions/ordnance, chemical engineering, or closely related disciplines) • Ability to evaluate the technical accuracy and real-world consequence of model outputs in your domain • Understanding of dual-use research concerns and the distinction between open-source knowledge and operationally significant uplift • Strong hands-on experience using multiple LLMs (ChatGPT, Claude, Gemini, open-source models, etc.) • Creative, adversarial problem-solving skills • Clear and precise written communication, including the ability to explain technical risk to non-specialist audiences • Strong ethical judgment and the ability to separate adversarial thinking from personal values • Self-directed, collaborative, and comfortable in feedback-heavy environments • Candidates must be able to engage with sensitive CBRNE-related material professionally and sustainably • Active or prior security clearance (Secret, Top Secret, or SCI) (nice to have) • Experience in threat assessment, WMD analysis, intelligence analysis, or arms control verification (nice to have) • Background in biosafety/biosecurity, chemical safety, nuclear nonproliferation, or explosive ordnance disposal (nice to have) • Familiarity with relevant regulatory frameworks including CWC, BWC, IAEA safeguards, ATF regulations, and Export Administration Regulations (nice to have) • Experience in red teaming, penetration testing, or structured adversarial evaluation (nice to have) • Familiarity with Python or scripting languages, LLM APIs, or evaluation tooling (nice to have) • Published research or professional presentations in a relevant CBRNE domain (nice to have) • Prior work in trust and safety, content moderation, or AI evaluation (nice to have)
Apply Nowđź•’ Yesterday
51 - 200
AI safety red teamer probing conversational models for jailbreaks, prompt injections, and systemic risks. Producing human-data artifacts that help Mercor customers build safer frontier AI.
🇺🇸 United States – Remote
đź’µ $29 - $45 / hour
🔥 Funding within the last year
đź’° $350M Series C - Mercor on 2025-10
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
🗣️🇧🇷🇵🇹 Portuguese Required
đź•’ Yesterday
51 - 200
AI Safety Red Teamer stress-testing frontier AI models for Mercor. Identifying vulnerabilities across high-risk domains and improving model alignment, robustness, and safety.
🇺🇸 United States – Remote
đź’µ $70 - $84 / hour
🔥 Funding within the last year
đź’° $350M Series C - Mercor on 2025-10
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
đź•’ Yesterday
51 - 200
AI Safety Red Teamer identifying vulnerabilities in frontier AI models for Mercor’s AI lab partnerships. Designing adversarial prompts and evaluating high-risk model behavior.
🇺🇸 United States – Remote
đź’µ $70 - $84 / hour
🔥 Funding within the last year
đź’° $350M Series C - Mercor on 2025-10
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
đź•’ Yesterday
51 - 200
AI Safety Red Teamer stress-testing frontier AI models for Mercor’s AI training partnerships. Identifying vulnerabilities and improving model alignment, robustness, and safety.
🇺🇸 United States – Remote
đź’µ $70 - $84 / hour
🔥 Funding within the last year
đź’° $350M Series C - Mercor on 2025-10
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence
đź•’ Yesterday
51 - 200
AI safety red team expert probing Mercor’s partner AI models for vulnerabilities. Generating reproducible adversarial data, reports, and attack cases to improve model safety.
🇺🇸 United States – Remote
đź’µ $20 - $22 / hour
🔥 Funding within the last year
đź’° $350M Series C - Mercor on 2025-10
⏳ Contract/Temporary
🟡 Mid-level
đźź Senior
🤖 Artificial Intelligence