AI Red Team Engineer

Job not on LinkedIn

🕒 July 1

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of White Circle

White Circle

1 - 10 employees

đŸ€– Artificial Intelligence

🔌 API

🔐 Security

Artificial Intelligence ‱ API ‱ Security

White Circle is a unified control layer for AI applications that helps teams test, protect, observe, and optimize model-driven features. The platform provides automated red-teaming, low-latency guardrails, input/output protection (PII detection, jailbreak prevention, hallucination detection, data-leak prevention), analytics and custom metrics (risk scoring, user-behavior clustering), and optimization tools like dynamic prompts and model routing. Delivered via a high-throughput API with enterprise-grade security (SOC 2, HIPAA), White Circle targets organizations deploying AI at scale.

📋 Description

‱ Red-team LLM-powered systems: chatbots, copilots, RAG pipelines, AI agents, tool-calling workflows, and API-based AI products. ‱ Test for jailbreaks, prompt injection, system-prompt and tool leakage, sensitive-data and context leakage, unsafe outputs, policy bypass, tool misuse, excessive agency, resource and token-cost abuse, and business-logic abuse. ‱ Write lightweight Python to automate attacks, run prompt sets, call model APIs, collect and score responses, and generate repeatable reports. ‱ Build and maintain an internal attack library: prompts, scenarios, test cases, regression tests, scoring rubrics, and reusable demo cases. ‱ Turn model failures into clear reports: what happened, why it matters, how to reproduce it, how severe it is, and how to fix it. ‱ Convert successful attacks into regression tests and product requirements. ‱ Track new red-team and safety techniques and fold the useful ones into our tests. ‱ Support GTM by producing strong, credible evidence for customer demos, security reviews, and sales conversations.

🎯 Requirements

‱ Genuinely love breaking things and reasoning adversarially. ‱ Have a background in QA automation, AppSec, API/security/pen testing, or bug bounty. ‱ Have strong Python scripting skills. ‱ Have experience testing APIs, web apps, backends, or SaaS products. ‱ Are hands-on with LLMs, prompts, system instructions, RAG, agents, and tool/function calling. ‱ Understand LLM-specific abuse vectors (prompt injection, jailbreaks, data leakage, tool misuse, excessive agency, token-cost exhaustion). ‱ Can find bypasses, abuse edge cases, chain failures, and reason about real-world impact. ‱ Can separate real customer risk from low-impact prompt tricks. ‱ Write clear, reproducible bug reports in clear English. ‱ Can move fast without perfect requirements. ‱ Hold a firm ethical line: you red-team to make systems safer, operate within scope and the law, and don't produce or traffic in genuinely harmful material.

đŸ–ïž Benefits

‱ Paid time off in line with your local regulations, no matter where you work from ‱ Work from Paris (hybrid) + relocation package ‱ Best medical insurance in France ‱ All the hardware, tools, and services you need ‱ Covered subscriptions for AI agents ‱ Team off-sites twice a year: we've recently been to the Alps and to Saint-Tropez

Apply Now

Similar Jobs

🕒 March 17

Iris.ai

11 - 50

đŸ’Œ Consulting

đŸ€– Artificial Intelligence

☁ SaaS

Senior AI Sales Executive driving growth by managing complex sales cycles across AI and data sectors for Iris.ai. Engaging high-potential enterprise clients and supporting strategic accounts.

🕒 March 14

Velenosi & Meredith Consulting

1 - 10

đŸ€ B2B

đŸ’Œ Consulting

🎯 Recruiter

Dutch-speaking Prompt Engineer involved in AI-powered clinical documentation workflows. Developing prompts for language models impacting healthcare documentation across Europe.

đŸ—ŁïžđŸ‡łđŸ‡± Dutch Required