AI Safety Specialist, English – Danish

Job not on LinkedIn

🔥 0 minutes ago

🗣️🇩🇰 Danish Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of 24-MAG

24-MAG

2 - 10 employees

🤝 B2B

💼 Consulting

B2B • Consulting

24-MAG is a commercial strategy and execution firm that helps B2B organizations design and implement systems, workflows, and operating rhythms for sales, client management, and cross-functional projects. They focus on transforming scattered processes into aligned, measurable, and scalable commercial functions—covering pipeline structure, account management frameworks, and operational discipline for teams seeking efficient, intentional growth.

📋 Description

• Conduct adversarial testing of conversational AI models and agents • Develop jailbreaks, prompt-injection scenarios, misuse cases, and multi-turn manipulation strategies • Probe models for vulnerabilities across diverse conversational and adversarial scenarios • Apply systematic testing frameworks • Identify and classify model failures and safety vulnerabilities • Evaluate bias, misinformation, misuse, and potentially harmful model behaviours • Assess vulnerability severity, reproducibility, and practical significance • Produce high-quality human evaluation data from red teaming activities • Annotate model failures and categorise identified vulnerabilities • Create structured attack cases and evaluation materials • Document adversarial scenarios clearly and reproducibly • Produce reports, datasets, and structured findings for technical teams • Explain identified risks to technical and non-technical stakeholders • Record testing methodology, model behaviour, and failure patterns • Contribute to broader evaluation coverage across models and use cases

🎯 Requirements

• Native-level fluency in both English and Danish • Prior experience with AI red teaming, adversarial AI evaluation, cybersecurity, or socio-technical system testing • Strong understanding of conversational AI systems and model failure modes • Experience developing structured adversarial tests and evaluation frameworks • Ability to identify subtle vulnerabilities and recurring behavioural patterns • Strong analytical reasoning and written communication skills • Ability to document findings clearly and reproducibly • Comfort working across changing projects, scenarios, and evaluation frameworks • Background in computer science, cybersecurity, artificial intelligence, machine learning, linguistics, behavioural science, or a related discipline may be helpful • Equivalent professional experience in AI safety, adversarial testing, security, or structured model evaluation may also be considered • Experience creating jailbreak or prompt-injection datasets • Familiarity with adversarial machine learning • Knowledge of RLHF, DPO, model extraction, or related AI training and evaluation concepts • Cybersecurity experience involving penetration testing, exploit development, or reverse engineering • Background analysing abuse, harassment, misinformation, or other socio-technical risks • Experience testing conversational AI systems • Previous experience producing structured human data for AI evaluation • Native-level English and Danish proficiency is required • Independent contractor status

🏖️ Benefits

• Flexible scheduling • Competitive hourly compensation ($45–$60 per hour) • Weekly payments via Stripe or Wise • Optional participation in higher-sensitivity projects • Flexible remote consulting work • Projects may be extended, shortened, or adjusted depending on scope and performance

Apply Now

Similar Jobs

🔥 1 hour ago

Lightly

11 - 50

🤖 Artificial Intelligence

🤝 B2B

☁️ SaaS

Energy sector expert evaluating AI-generated forecasts for Lightly AG, an ETH and HSG spin-off building machine-learning and computer-vision technology. Assessing forecast accuracy and response quality using a provided rubric.

🔥 18 hours ago

Cayuse Holdings

501 - 1000

💼 Consulting

📦 Logistics

🏥 Healthcare

AI Native Engineer building Claude Code agents, skills, and MCP integrations. Developing AWS-backed harness components for retrieval, tools, evaluation, and observability.

🕒 Yesterday

Terac

1 - 10

🤖 Artificial Intelligence

🤝 B2B

STEM experts creating and solving PhD-level academic problems for Terac’s paid AI reasoning research studies. Reviewing AI-generated technical responses for logical and factual errors.

🕒 Yesterday

Terac

1 - 10

🤖 Artificial Intelligence

🤝 B2B

STEM experts creating graduate-to-PhD-level problems and exact solutions for Terac’s AI training research. Reviewing AI responses and identifying scientific reasoning errors.

🕒 Yesterday

Hippocratic AI

11 - 50

🤖 Artificial Intelligence

⚕️ Healthcare Insurance

🏥 Healthcare

Evaluating Hippocratic AI’s clinical agents for symptom triage, diabetes medication titration, and travel health guidance. Providing safety-focused feedback to improve AI-driven healthcare tools.