AIPayList

Jobs / Terac

AI Evaluators: Assessing a Shopping Assistant

$40/hr

source wording: “Terac listing record, payRate 40.00 USD HOURLY

What this role asks for

  • AI Evaluation
  • Quality Assurance
  • E-Commerce
  • Data Annotation
  • Prompt engineering
  • Annotation
Read from the posting's own words: show the exact sentences
  • AI Evaluation: “listed by the platform under required skills: AI Evaluation
  • Quality Assurance: “listed by the platform under required skills: Quality Assurance
  • E-Commerce: “listed by the platform under required skills: E-Commerce
  • Data Annotation: “listed by the platform under required skills: Data Annotation
  • Prompt engineering: “…eye for detail. We welcome applicants with prior experience in prompt engineering, complex data annotation, or software testing. You should be comfortable analyzing text…
  • Annotation: “…applicants with prior experience in prompt engineering, complex data annotation, or software testing. You should be comfortable analyzing text interactions deeply and…

The role, as Terac describes it

We are running a paid project to evaluate the performance of an AI shopping assistant. We need detail-oriented reviewers to analyze real interaction traces, identify failures, and build quality rubrics.

What We're Researching

We're hiring AI evaluators to assess the accuracy and helpfulness of a new digital shopping assistant. This project focuses on understanding how well the system handles real-world e-commerce queries and where it falls short in its logic. Your analysis will directly feed into improving the underlying model and its response quality. How It Works

Read the rest of the description (12 more paragraphs)

You will review real interaction traces between users and the shopping assistant within our custom platform. As you analyze these conversations, you will pinpoint specific failures, logical errors, or unhelpful product recommendations. From there, you will create structured rubrics and verifiers to consistently judge future response quality. This is an ongoing remote engagement requiring 20+ hours per week. Who This Is For

This opportunity is ideal for quality assurance specialists, AI data evaluators, and e-commerce professionals with a strong eye for detail. We welcome applicants with prior experience in prompt engineering, complex data annotation, or software testing. You should be comfortable analyzing text interactions deeply and building structured evaluation frameworks from scratch.

What you would do

• Review real user interaction traces with an AI shopping assistant

• Identify logical failures, inaccuracies, or poor recommendations in the text

• Create structured rubrics and verifiers to judge response quality

• Commit to 20+ hours per week of evaluation work on our internal platform

Who this is for

• Experience in data evaluation, quality assurance, or AI training

• Strong analytical skills with the ability to spot subtle errors in text

• Familiarity with e-commerce search and digital shopping experiences

• Ability to commit to a sustained workload of 20+ hours per week

Posted by Terac, reproduced here so you can judge the role before clicking. Original posting ↗

Source: platform job feed · first seen 4d ago · ID TKd76RUIQxw3wy-x

More like this

$40/hrTerac · AI Evaluators: Assessing a Shopping Assistant

Apply on Terac