Video Games Reviewer – AI Evaluation

🔥 0 minutes ago

🇺🇸 United States – Remote

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

👻 Ghost score 12%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Gramian Consulting

Gramian Consulting

2 - 10 employees

Founded 2025

💼 Consulting

📦 Logistics

📣 Marketing

Consulting • Logistics • Marketing

Gramian Consulting is a remote-first consulting firm that connects engineering and data/AI talent with organizations through talent augmentation, recruiting, dedicated teams, and contractor management. The firm provides Data & AI services including LLM training and fine-tuning, AI agents and assistants, MLOps, and AI infrastructure, and it offers mentorship and education programs for career readiness, interview preparation, and international market orientation. Rooted in hands-on engineering and recruiting experience, Gramian helps clients scale technical teams and extract business value from AI while developing individual talent.

📋 Description

• Review and validate video game-focused prompts across a broad range of gaming topics • Evaluate completed tasks for factual accuracy, reasoning quality, completeness, and relevance • Identify factual errors, hallucinations, logical inconsistencies, and weak annotations • Verify claims related to games, publishers, platforms, genres, esports, and game development • Assess whether prompts are sufficiently challenging and aligned with project objectives • Ensure compliance with project instructions, review rubrics, and quality standards • Provide clear, constructive, and evidence-based feedback to contributors • Document review findings and escalate ambiguous or complex cases when required • Collaborate with project managers and AI teams to improve evaluation quality and workflows

🎯 Requirements

• Master’s degree or higher in Game Design, Interactive Media, Computer Science, or a closely related discipline • At least 3 years of professional, research, teaching, or industry experience in the video games domain • Strong knowledge of game development, gaming platforms, publishers, genres, esports, or gaming communities • Experience reviewing, evaluating, editing, or quality-checking technical or domain-specific content • Ability to identify subtle factual, logical, and contextual errors • Excellent written English communication skills • Strong analytical ability and attention to detail • Ability to work independently in a quality-focused remote environment

Apply Now

Similar Jobs

🔥 20 hours ago

Mercor

51 - 200

Lifecycle marketing expert designing simulated environments to evaluate AI agents on real CRM and retention work. Creating briefs, rubrics, and reference answers for Mercor’s AI training projects.

🇺🇸 United States – Remote

💵 $60 - $100 / hour

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

🔥 20 hours ago

Mercor

51 - 200

Field marketing expert designing realistic task environments for Mercor’s AI training projects. Evaluating agent performance across events, pipeline recovery, sponsorships, and multi-tool workflows.

🇺🇸 United States – Remote

💵 $60 - $100 / hour

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

🕒 Yesterday

The Browser Company

11 - 50

💼 Consulting

📣 Marketing

AI Prototyper exploring frontier AI for Dia, The Browser Company’s next-generation browser. Building and evaluating prototypes that guide product strategy and engineering handoff.

🇺🇸 United States – Remote

💵 $100 - $150 / hour

💰 $13M Series A on 2021-05

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence

🕒 Yesterday

Terac

1 - 10

🤖 Artificial Intelligence

🤝 B2B

Terac, which recruits vetted human experts for AI research, runs a paid study on designing challenging prompts. Participants test ChatGPT failures and create grading rubrics.

🕒 Yesterday

Mercor

51 - 200

AI safety red teamer probing Mercor-partner AI models with adversarial attacks. Generating reproducible datasets and reports to strengthen frontier AI systems.

🇺🇸 United States – Remote

💵 $20 - $22 / hour

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 Artificial Intelligence