TypeScript Engineer, AI Coding Agent Evaluator

Job not on LinkedIn

🔥 15 hours ago

🇺🇸 United States – Remote

💵 $100 - $200 / hour

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 AI Engineer

👻 Ghost score 7%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of G2i Inc.

G2i Inc.

11 - 50 employees

🎯 Recruiter

🏢 Enterprise

☁️ SaaS

Recruitment • Enterprise • SaaS

G2i Inc. is a video-based platform specializing in the hiring of contract or full-time engineers. With a vast pool of over 8,000 engineers, G2i offers quality matches quickly by customizing its sourcing process based on client needs. They provide a 7-day free trial to engage potential hires before making a commitment. The platform emphasizes speed and quality, utilizing video assessments and AI training to ensure engineers meet specific technical standards. G2i handles compliance, international payments, and background checks, providing a streamlined hiring process for companies looking to hire in the US, Canada, Latin America, or Europe. By focusing on technical roles such as JavaScript, Python, iOS, Android, and AI engineers, G2i ensures companies can hire top talent efficiently.

📋 Description

• Evaluate AI-generated coding interactions end-to-end • Judge whether responses are useful, correct at a high level, and aligned with strong engineering thinking • Assess the quality of explanations, preambles, and reasoning—not just code • Distinguish different levels of response quality • Provide clear, opinionated feedback on what worked, what failed, and what felt misleading • Help define what great interaction looks like for tools such as Cursor • Make subjective but rigorous judgments about AI coding-agent behavior

🎯 Requirements

• Highly experienced software engineer (SR+) or Staff/Principal-level engineer (or equivalent experience) • Strong background in TypeScript/JavaScript or Python • Hands-on experience using OpenAI Codex, Claude Code, and Cursor • Deep familiarity with modern AI-assisted development workflows • Ability to evaluate code without fully executing or deeply reviewing every line • Comfortable providing direct, opinionated feedback • High standards for engineering quality • Must complete a take-home evaluation exercise and one behavioral interview • Must agree to a simple background check for the project

Apply Now

Similar Jobs

🕒 2 days ago

Mercor

51 - 200

AI evaluator reviewing software, IT, and data artifacts for accuracy and quality. Providing structured feedback to help Mercor and leading AI labs train frontier models.

🇺🇸 United States – Remote

💵 $80 - $120 / hour

🔥 Funding within the last year

💰 $350M Series C - Mercor on 2025-10

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🤖 AI Engineer

🕒 6 days ago

Accellor

201 - 500

💼 Consulting

🏥 Healthcare

🏨 Hospitality

Senior Applied AI Engineer building production Python and Azure solutions for a global media and entertainment organization. Developing LLM, RAG, backend, and agentic AI applications.

🕒 6 days ago

9th Way Insignia

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

AI Engineer designing secure production AI/ML and LLM systems for the Department of Veterans Affairs. Evaluating models, managing risks, and delivering responsible AI capabilities.

🕒 August 28

Arctiq

201 - 500

💼 Consulting

🏥 Healthcare

📦 Logistics

Automation & AI Engineer modernizing a client’s ITSM operations for Arctiq. Building AI-powered workflows, integrations, self-service, predictive operations, and self-healing capabilities.

🕒 August 17

Brain Gain Recruiting

1 - 10

🎯 Recruiter

💼 Consulting

☁️ SaaS

Explainable AI Engineer developing validated prediction analytics and scalable AI pipelines. Supporting aerospace and defense supply-chain decisions with explainable, mission-critical Agentic AI.