OpenRecently verified

AI Evaluators: Assessing A Shopping Assistant

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

$50/hrAs published by the platform.

About this job

This project asks evaluators to examine how a digital shopping assistant answers real e-commerce questions and to locate where its reasoning or product suggestions fall short. The work is remote and ongoing, and it is aimed at people with quality assurance, data evaluation, or e-commerce backgrounds who can build structured evaluation frameworks.

What you'll do

  • Review recorded conversations between users and an AI shopping assistant.
  • Identify logical errors, inaccuracies, and poor product recommendations in the responses.
  • Build structured rubrics and verifiers to judge the quality of future responses.

Requirements

Quality assuranceData evaluationAI trainingE-commerce searchSoftware testingPrompt engineeringData annotation

    Pay

    $50/hr

    Pay as published by the platform. It is not a guarantee of income or hours.

    Location

    Open to residents of: United States.

    How to apply

    You apply on Terac's posting: start its screening from the link in the listing. If you qualify, Terac invites you to the paid tasks.

    All Terac jobs and how the platform works

    Source

    Official posting: https://jobs.ashbyhq.com/terac/7f03944b-faf5-4829-b40c-2f9edf2fe0ec

    Last checked on October 8, 2026.

    Similar jobs

    • Data & Machine Learning
    • Platform: Alignerr

    Review and evaluate AI-generated outputs across text, images, and structured data to assess quality, accuracy, and consistency. This remote contract role is ideal for detail-oriented individuals who can identify errors and provide actionable feedback to guide AI model improvement, with no prior AI experience required.

    • Data & Machine Learning
    • Platform: Alignerr

    This role involves completing structured evaluation and labeling tasks to train and improve AI systems. You will assess AI-generated content across multiple formats and provide feedback to help make AI systems more accurate and reliable, working remotely on a flexible schedule.

    • Data & Machine Learning
    • Platform: micro1
    • Location: 58 countries
    • Level: Intermediate

    This role involves evaluating and monitoring the quality of datasets intended to train next-generation AI systems. You will examine detailed outputs to verify their accuracy, completeness and compliance with guidelines, applying structured evaluation grids and documenting your analyses. The ideal profile masters data review and possesses solid domain expertise, though AI experience is not required.

    Posted September 27, 2026

    OpenRecently verified$30-60/hr

    Referral link to micro1's job list: search for this role there. Open this exact job

    • Data & Machine Learning
    • Platform: Mercor
    • Location: 40 countries
    • Level: Expert

    This role involves judging how frontier AI models respond to sensitive and ambiguous topics, checking answers for safety, accuracy and policy compliance. It suits experienced professionals from fields such as psychology, law, science or public policy who can apply safety rubrics and give structured feedback to model developers.

    Posted July 16, 2026

    OpenRecently verified$60-70/hr

    Referral link: we may earn a fee. Apply without it

    • Data & Machine Learning
    • Platform: Mercor
    • Location: United States
    • Language: Bangla

    This role involves testing conversational AI models as an AI safety expert by generating adversarial training data to identify vulnerabilities. You will produce reports and structured datasets that enable clients to improve the robustness and security of their AI systems. The ideal profile masters critical content analysis, methodological rigor, and technical communication.

    Posted June 4, 2026

    OpenRecently verified$16-22/hr

    Referral link: we may earn a fee. Apply without it

    • Data & Machine Learning
    • Platform: Vetto

    Vetto is hiring quality assurance reviewers to audit responses and tasks produced by human annotators for AI projects. The work checks that delivered data meets high quality standards and flags content that appears AI-generated or departs from the project guidelines. It suits people who can judge data accuracy and consistency against detailed evaluation criteria.

    Posted March 4, 2026

    OpenRecently verifiedPay not disclosed

    Keep exploring