Search Remote Jobs

AI Safety Specialist, English, Norwegian

Job not on LinkedIn

šŸ”„ 1 minute ago

šŸ—£ļøšŸ‡³šŸ‡“ Norwegian Required

Apply Now
Find Similar Remote Jobs

šŸ“Š Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of 24-MAG

24-MAG

2 - 10 employees

šŸ¤ B2B

šŸ’¼ Consulting

B2B • Consulting

24-MAG is a commercial strategy and execution firm that helps B2B organizations design and implement systems, workflows, and operating rhythms for sales, client management, and cross-functional projects. They focus on transforming scattered processes into aligned, measurable, and scalable commercial functions—covering pipeline structure, account management frameworks, and operational discipline for teams seeking efficient, intentional growth.

šŸ“‹ Description

• Conduct adversarial testing of conversational AI models and agents • Develop jailbreaks, prompt-injection scenarios, misuse cases, and multi-turn manipulation strategies • Probe models for vulnerabilities across diverse conversational and adversarial scenarios • Identify and classify model failures and safety vulnerabilities • Evaluate bias, misinformation, misuse, and potentially harmful model behaviours • Assess vulnerability severity, reproducibility, and practical significance • Produce and annotate human evaluation data from red teaming activities • Create structured attack cases and evaluation materials • Document adversarial scenarios, testing methodology, model behaviour, and failure patterns • Produce reports, datasets, and structured findings for technical teams • Explain identified risks to technical and non-technical stakeholders • Contribute to broader evaluation coverage across models and use cases

šŸŽÆ Requirements

• Native-level fluency in both English and Norwegian • Prior experience with AI red teaming, adversarial AI evaluation, cybersecurity, or socio-technical system testing • Strong understanding of conversational AI systems and model failure modes • Experience developing structured adversarial tests • Ability to identify subtle vulnerabilities and recurring behavioural patterns • Strong analytical reasoning and written communication skills • Ability to document findings clearly and reproducibly • Comfort working across changing projects, scenarios, and evaluation frameworks • Equivalent professional experience in AI safety, adversarial testing, security, or structured model evaluation may be considered • Practical red teaming experience is particularly valuable • Familiarity with adversarial machine learning • Knowledge of RLHF, DPO, model extraction, or related AI training and evaluation concepts • Cybersecurity experience involving penetration testing, exploit development, or reverse engineering • Experience analysing abuse, harassment, misinformation, or other socio-technical risks • Experience testing conversational AI systems • Previous experience producing structured human data for AI evaluation

šŸ–ļø Benefits

• Fully remote work • Flexible scheduling • Competitive hourly compensation • Weekly payments via Stripe or Wise • Participation in higher-sensitivity projects is optional • Flexible remote consulting work • Projects may be extended, shortened, or adjusted depending on scope and performance

Apply Now

Similar Jobs

šŸ”„ 1 hour ago

Lightly

11 - 50

šŸ¤– Artificial Intelligence

šŸ¤ B2B

ā˜ļø SaaS

Energy sector expert evaluating AI-generated forecasts for Lightly AG, an ETH and HSG spin-off building machine-learning and computer-vision technology. Assessing forecast accuracy and response quality using a provided rubric.

šŸ”„ 18 hours ago

Cayuse Holdings

501 - 1000

šŸ’¼ Consulting

šŸ“¦ Logistics

šŸ„ Healthcare

AI Native Engineer building Claude Code agents, skills, and MCP integrations. Developing AWS-backed harness components for retrieval, tools, evaluation, and observability.

šŸ•’ Yesterday

Terac

1 - 10

šŸ¤– Artificial Intelligence

šŸ¤ B2B

STEM experts creating and solving PhD-level academic problems for Terac’s paid AI reasoning research studies. Reviewing AI-generated technical responses for logical and factual errors.

šŸ•’ Yesterday

Terac

1 - 10

šŸ¤– Artificial Intelligence

šŸ¤ B2B

STEM experts creating graduate-to-PhD-level problems and exact solutions for Terac’s AI training research. Reviewing AI responses and identifying scientific reasoning errors.

šŸ•’ Yesterday

Hippocratic AI

11 - 50

šŸ¤– Artificial Intelligence

āš•ļø Healthcare Insurance

šŸ„ Healthcare

Evaluating Hippocratic AI’s clinical agents for symptom triage, diabetes medication titration, and travel health guidance. Providing safety-focused feedback to improve AI-driven healthcare tools.