AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

682
open jobs
29
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 640 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 9, 2026.

441 to 460 of 682 jobs
  • Science & Research
  • Platform: Mercor
  • Location: United States

Mercor is looking for experienced biotechnologists to evaluate and rate the work of AI models on biotechnology research and development tasks. You will provide expert evaluations on scientific validity, development feasibility, and regulatory compliance of proposals generated by AI, drawing on your hands-on experience with completed biotechnology projects.

Posted September 15, 2026

OpenRecently verified$110-125/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States

You will put your experience in operational management to the service of training and evaluating AI models. The work consists of writing real cases from your expertise, evaluating operational recommendations generated by AI, and identifying gaps between what seems viable theoretically and what works in practice. This role is aimed at experienced operations professionals capable of structuring their business judgment.

Posted September 15, 2026

OpenRecently verified$90-120/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

Join a network of chemistry experts to evaluate and train the most advanced AI models. You will contribute to improving AI systems by assessing the quality of their chemical proposals, explaining your technical expertise, and flagging discrepancies between theory and practice.

Posted September 15, 2026

OpenRecently verified$65-105/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States

Mercor is seeking consumer and lifestyle experts to evaluate and improve AI model responses on real-world everyday questions. You will write reference questions, rate model recommendations based on their relevance and safety, and flag dangerously inaccurate answers. This role is suitable for professionals with recognized expertise in a specific consumer category (sports, film, gastronomy, fashion, consumer electronics, etc.).

Posted September 15, 2026

OpenRecently verified$50-75/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States

You will join a network of data experts to assess and evaluate the quality of analyses produced by frontier AI models. Your roles will consist of designing analysis problems, evaluating pipelines and queries generated by the models, and identifying results that are technically valid but potentially misleading for decision-makers. This role is aimed at data practitioners with concrete experience in delivered analytical projects.

Posted September 15, 2026

OpenRecently verified$70-120/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States

Join a network of education experts to evaluate and improve explanations generated by AI models. You will write pedagogical exercises, rate materials created by AI, and identify points of confusion, leveraging your experience as a teacher or trainer.

Posted September 15, 2026

OpenRecently verified$30-85/hr

Referral link: we may earn a fee. Apply without it

  • Engineering
  • Platform: Mercor
  • Location: United States

This role consists of evaluating and rating AI model responses to electrical engineering problems by applying your technical expertise and professional standards. You will participate in specific projects that require either writing problem statements with their reference solutions or verifying the technical relevance of AI-generated designs. This work is intended for experienced electrical engineers in design, testing, or systems who are able to communicate their technical reasoning clearly.

Posted September 15, 2026

OpenRecently verified$70-120/hr

Referral link: we may earn a fee. Apply without it

  • Engineering
  • Platform: Mercor
  • Location: United States

Mercor is looking for experienced engineers to evaluate and rate technical solutions proposed by AI models in your area of expertise. You will critique the quality of analyses and designs generated, identifying gaps between theory and practical feasibility. This role is open on a continuous basis for engineers interested in AI evaluation and training work.

Posted September 15, 2026

OpenRecently verified$60-125/hr

Referral link: we may earn a fee. Apply without it

  • Finance & Accounting
  • Platform: Mercor
  • Location: United States

This role consists of evaluating and rating financial work produced by AI models, applying your professional expertise as a benchmark to judge the quality of their analyses. You will contribute to training AI systems by validating accounting rigor, valuation assumptions, and regulatory compliance of their results.

Posted September 15, 2026

OpenRecently verified$60-180/hr

Referral link: we may earn a fee. Apply without it

  • Engineering
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

You design assessment exercises to train and test AI models specialized in fire safety and occupant protection. Starting from real plans, documents and site photos, you create realistic tasks (evacuation analyses, correction lists, control decisions) and validate the answers with an expert peer before distribution.

Posted September 15, 2026

OpenRecently verified$45-60/hr

Referral link: we may earn a fee. Apply without it

  • Generalist & Data Labeling
  • Platform: Mercor
  • Location: United States

This role involves participating in the evaluation and improvement of AI models by providing expert human judgment. You will be asked to write reference questions, rate generated responses, and identify errors or inconsistencies in models. The position suits candidates from varied educational backgrounds who master close reading and written communication.

Posted September 15, 2026

OpenRecently verified$40-70/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States

Mercor is looking for humanities researchers to evaluate and rate responses generated by AI models in their respective fields of expertise. The work involves creating reference questions, assessing the quality of interpretations and arguments produced by AI, and flagging reading errors or hallucinations. This role is suited for humanities specialists who wish to participate in AI development through evaluation.

Posted September 15, 2026

OpenRecently verified$45-60/hr

Referral link: we may earn a fee. Apply without it

  • Finance & Accounting
  • Platform: Mercor
  • Location: United States

You put your investment banking expertise to work training and evaluating generative AI. You create evaluation scorecards, review the outputs of models in financial modeling and deal materials, and explain deviations from industry standards.

Posted September 15, 2026

OpenRecently verified$100-130/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: Mercor
  • Location: United States
  • Language: Spanish, Arabic, Hindi, French, Portuguese

This role consists of evaluating and improving the capabilities of AI models in language and audio. You will provide expert judgments on the quality of translations, the naturalness of voice and audio content, and participate in the creation of training resources. This role is designed for linguists, translators, and audio professionals with recognized native or professional expertise.

Posted September 15, 2026

OpenRecently verified$35-50/hr

Referral link: we may earn a fee. Apply without it

  • Legal
  • Platform: Mercor
  • Location: United States

You participate in the training and evaluation of AI models by bringing your legal expertise to bear on improving these systems. You analyze solutions generated by AI, you evaluate their compliance with applicable law, and you write reference legal cases. This role suits legal professionals, regulatory compliance experts, or public affairs specialists who master written communication and know how to navigate ambiguous contexts.

Posted September 15, 2026

OpenRecently verified$60-150/hr

Referral link: we may earn a fee. Apply without it

  • Legal
  • Platform: Mercor
  • Location: United States

You will evaluate and rate legal work produced by AI models for cutting-edge artificial intelligence research teams. Your tasks will consist of drafting legal motions from factual cases, reviewing contracts and discovery documents generated by AI, and verifying the legal accuracy and relevance of citations. This role is aimed at experienced lawyers in litigation, transactional law, or regulatory law who can precisely justify their reasoning and are comfortable with ambiguous work statements.

Posted September 15, 2026

OpenRecently verified$100-150/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

Mercor is looking for life, physical and social science researchers to evaluate and rate work generated by AI models. Your role will involve designing research problems, rating the analysis produced by AI systems, and identifying methodological gaps, using your scientific expertise to establish evaluation criteria.

Posted September 15, 2026

OpenRecently verified$60-120/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States

This role involves evaluating and judging AI models as a machine learning expert, providing the critical expertise that research teams cannot generate alone. You will use your practical experience to assess modeling problems, training code, and production approaches, explaining your reasoning clearly. This permanent roster recruits machine learning practitioners ready to participate in AI training and evaluation projects.

Posted September 15, 2026

OpenRecently verified$70-120/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States

Consulting role where you will evaluate and assess the capabilities of AI models on real-world strategy, operations, or transformation cases. You will write case statements based on your experience, evaluate market analyses and recommendations generated by AI, and flag answers that are overly polished to withstand rigorous scrutiny. Intended for experienced consultants who wish to contribute to AI training.

Posted September 15, 2026

OpenRecently verified$90-120/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

This role involves evaluating and rating responses generated by AI models in materials science, using your expertise to judge the quality of their reasoning. You will be asked to formulate materials problems based on your previous work, assess the consistency between microstructure and properties, and identify cases where proposed materials fail under real-world conditions. This role is aimed at materials scientists with solid research or industrial experience and capable of clearly communicating their expertise.

Posted September 15, 2026

OpenRecently verified$60-120/hr

Referral link: we may earn a fee. Apply without it

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

682 AI Evaluation AI training jobs are open on SideHustler today. 29 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 9, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 640 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (142), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.