Search Remote Jobs

Researcher, Evaluations and Benchmarks

Job not on LinkedIn

🔥 0 minutes ago

🗽 New York – Remote

infoinfo

⏰ Full Time

🟡 Mid-level

đźź  Senior

đź‘» Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of ActiveFence

ActiveFence

201 - 500 employees

đź”’ Cybersecurity

🤖 Artificial Intelligence

đź’° $400M Series B on 2021-07

Cybersecurity • Artificial Intelligence

ActiveFence is a Trust and Safety provider for online platforms, protecting platforms and their users from malicious behavior and content. Trust and Safety teams of all sizes rely on ActiveFence to keep their users safe from the widest spectrum of online harms, unwanted content, and malicious behavior, including child safety and exploitation, disinformation, hate speech, terror, nudity, fraud, and more. We offer a full stack of capabilities with our deep intelligence research, AI-driven harmful content detection, and online content moderation platform. Protecting over three billion users globally everyday in over 100 languages, ActiveFence lets people interact and thrive online.

đź“‹ Description

• Ship a benchmark every two to three weeks measuring a previously unmeasured frontier risk • Collaborate with leading AI labs and universities on benchmarks and papers • Pair with in-house researchers who own specific harm areas and direct freelancers • Own the taxonomy, evaluation harness, quality bar, and release • Ensure frontier labs can rerun the benchmark and reproduce the reported numbers • Read evals, verify taxonomy alignment, and push researchers on quality • Own the benchmark plan and calendar and keep researchers on schedule • Direct two or three freelance subject matter experts as needed • Meet roughly monthly with the CTO, pod, and research leads to set the quarterly release roadmap • Tie the release plan to target accounts • Monitor the AI safety and security ecosystem, read research, maintain lab contacts, and gather feedback • Travel to conferences and speak with AI lab contacts weekly

🎯 Requirements

• PhD or Masters in computer science, machine learning or a related field, or equivalent depth from industry research • 3+ years building and running safety or security evaluations for language models in production, at an AI lab, a model provider, or a safety and security research organisation • 5+ relevant research publications in AI safety and security, including lead author on at least 2 • Strong engineering skills, including evaluation harnesses, distributed inference, vLLM, reading a codebase and fixing it • Ability to build a taxonomy, not only score against one • Ability to direct a researcher and two freelancers without formally managing them • Strong English, written and spoken • Curiosity about harms and ability to learn a new subject every three weeks • Post-training experience with SFT, DPO, or GRPO (ideally) • Reward design for subjective and safety-relevant targets (ideally) • Agentic evaluation experience with tool use, orchestration, permissions, and prompt injection (ideally) • Publications at top conferences (ideally) • Willingness to present work on client calls; strong verbal and written communication and ability to present to large and/or senior audiences (ideally)

🏖️ Benefits

• Budget for freelancers directed ad hoc when needed • Travel to conferences a couple of times a year • Travel to conferences at least 3 times a year (ideally)

Apply Now

Similar Jobs

đź•’ Yesterday

Twilio

5001 - 10000

🔌 API

🤝 B2B

Emerging Fraud Researcher detecting fraud across Twilio’s voice, messaging, and email network. Researching threats and developing AI-driven mitigation strategies with data science and engineering teams.

đź•’ September 8

VulnCheck

11 - 50

đź”’ Cybersecurity

🤖 Artificial Intelligence

🏢 Enterprise

Senior vulnerability researcher discovering and weaponizing vulnerabilities across firmware, software, and network devices. Advancing VulnCheck’s exploit intelligence and agentic vulnerability research.

đź•’ September 1

Zinpro Corporation

501 - 1000

🌾 Agriculture

🤝 B2B

Discovery Researcher developing ruminant nutrition products, research tools, and technical services at Zinpro, a global performance trace minerals and animal nutrition company.

🇺🇸 United States – Remote

đź’µ $105k - $175k / year

⏰ Full Time

🟡 Mid-level

đźź  Senior

đź•’ September 1

Zinpro Corporation

501 - 1000

🍽️ Food & Beverage

🏭 Manufacturing

🏥 Healthcare

Discovery Researcher advancing ruminant nutrition products, research, and digital tools. Supporting global technical sales teams for Zinpro’s science-led animal nutrition solutions.

đź•’ August 28

Bellese Technologies

51 - 200

🏥 Healthcare

đź’Ľ Consulting

⚕️ Healthcare Insurance

Researcher II conducting UX research for Bellese’s CMS Unified Case Management modernization. Improving Medicare and Medicaid program-integrity workflows through user-centered product insights.

🇺🇸 United States – Remote

đź’µ $109.4k - $142.8k / year

⏰ Full Time

🟡 Mid-level

đźź  Senior