The role, as Mercor describes it
Evaluate the quality and correctness of AI-assisted software-development traces used to train and evaluate a frontier AI lab's models. You'll assess end-to-end coding sessions produced with AI-assisted developer tools — judging correctness, workflow soundness, and reasoning — and provide clear, rubric-based written feedback.
Basic Qualifications
• 3+ years professional software development
• Hands-on experience with AI-assisted coding tools and agentic / spec-driven workflows (Cursor, GitHub Copilot, Claude Code, or similar)
• Strong code-reading and debugging skills across full-stack or backend systems
• Ability to evaluate multi-step coding trajectories for correctness and best practice
Preferred Qualifications
• Experience with Kiro or Amazon CodeCatalyst
Read the rest of the description (2 more paragraphs)
• Prior work evaluating or grading AI-generated code
• Contributions to developer tooling
Posted by Mercor, reproduced here so you can judge the role before clicking. Original posting ↗
Source: platform job feed · first seen 9d ago · ID list_AAABoEpz0r9yxMsCwGNKaJl4
More like this
- Apply ↗Meridial (Invisible)Coding & software evalWorldwideposted 3h ago
- Apply ↗ML Challenge Task Auditorup to $90/hrMercorCoding & software evalWorldwideposted 12d ago
- Apply ↗MercorCoding & software evalWorldwideposted 12d ago
- Apply ↗Meridial (Invisible)Coding & software evalWorldwideposted 3h ago
- Apply ↗SWE-Bench Task Auditorup to $90/hrMercorOther AI workWorldwideposted 12d ago
up to $90/hrMercor · AI Developer Trace Task Auditor
Apply on Mercor ↗