Search Remote Jobs

Researcher, Evaluations and Benchmarks

Job not on LinkedIn

🔥 0 minutes ago

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

đźź  Senior

đź‘» Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of ActiveFence

ActiveFence

201 - 500 employees

đź”’ Cybersecurity

🤖 Artificial Intelligence

đź’° $400M Series B on 2021-07

Cybersecurity • Artificial Intelligence

ActiveFence is a Trust and Safety provider for online platforms, protecting platforms and their users from malicious behavior and content. Trust and Safety teams of all sizes rely on ActiveFence to keep their users safe from the widest spectrum of online harms, unwanted content, and malicious behavior, including child safety and exploitation, disinformation, hate speech, terror, nudity, fraud, and more. We offer a full stack of capabilities with our deep intelligence research, AI-driven harmful content detection, and online content moderation platform. Protecting over three billion users globally everyday in over 100 languages, ActiveFence lets people interact and thrive online.

đź“‹ Description

• Ship a benchmark every two to three weeks measuring a previously unmeasured frontier risk • Collaborate with leading AI labs and universities on benchmarks and papers • Own the taxonomy, evaluation harness, quality bar, and release for each benchmark • Review evals personally and ensure verifiers, rubrics, distributions, and taxonomies meet quality standards • Hold the benchmark plan and calendar and keep researchers on timeline • Direct two or three ad-hoc SME freelancers • Set the quarterly release roadmap with the CTO, pod, and research leads based on research, client requests, and current events • Read relevant research and maintain relationships with AI labs • Speak with lab contacts weekly and attend conferences

🎯 Requirements

• PhD or Masters in computer science, machine learning or a related field, or equivalent depth from industry research • 3+ years building and running safety or security evaluations for language models in production, at an AI lab, a model provider, or a safety and security research organisation • 5+ relevant research publications in AI safety and security, including lead author on at least 2 • Strong engineering skills, including evaluation harnesses, distributed inference, vLLM, reading a codebase and fixing it • Ability to build a taxonomy, not only score against one • Ability to direct a researcher and two freelancers without formally managing them • Strong English, written and spoken • Curiosity about AI harms and ability to learn a new subject every three weeks • Ideally: post-training experience with SFT, DPO, and GRPO • Ideally: reward design for subjective and safety-relevant targets • Ideally: agentic evaluation experience with tool use, orchestration, permissions, and prompt injection • Ideally: publications at top conferences • Willingness to present work on client calls • Strong verbal and written communication and ability to present to large and/or senior audiences • Travel to conferences at least 3 times a year

🏖️ Benefits

• Travel to a couple of conferences a year • Travel to conferences at least 3 times a year • Opportunity to collaborate with leading AI labs and universities • Opportunity to work with approximately 150 researchers on AI harms • Role in the CTO office with access to research teams and freelancers

Apply Now

Similar Jobs

đź•’ Yesterday

Twilio

5001 - 10000

🔌 API

🤝 B2B

Emerging Fraud Researcher detecting fraud across Twilio’s voice, messaging, and email network. Researching threats and developing AI-driven mitigation strategies with data science and engineering teams.

đź•’ September 8

VulnCheck

11 - 50

đź”’ Cybersecurity

🤖 Artificial Intelligence

🏢 Enterprise

Senior vulnerability researcher discovering and weaponizing vulnerabilities across firmware, software, and network devices. Advancing VulnCheck’s exploit intelligence and agentic vulnerability research.

đź•’ September 1

Zinpro Corporation

501 - 1000

🌾 Agriculture

🤝 B2B

Discovery Researcher developing ruminant nutrition products, research tools, and technical services at Zinpro, a global performance trace minerals and animal nutrition company.

🇺🇸 United States – Remote

đź’µ $105k - $175k / year

⏰ Full Time

🟡 Mid-level

đźź  Senior

đź•’ September 1

Zinpro Corporation

501 - 1000

🍽️ Food & Beverage

🏭 Manufacturing

🏥 Healthcare

Discovery Researcher advancing ruminant nutrition products, research, and digital tools. Supporting global technical sales teams for Zinpro’s science-led animal nutrition solutions.

đź•’ August 28

Bellese Technologies

51 - 200

🏥 Healthcare

đź’Ľ Consulting

⚕️ Healthcare Insurance

Researcher II conducting UX research for Bellese’s CMS Unified Case Management modernization. Improving Medicare and Medicaid program-integrity workflows through user-centered product insights.

🇺🇸 United States – Remote

đź’µ $109.4k - $142.8k / year

⏰ Full Time

🟡 Mid-level

đźź  Senior