AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

682
open jobs
29
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 640 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 9, 2026.

301 to 320 of 682 jobs
  • Health & Medicine
  • Platform: Mercor
  • Location: United States

This role asks practicing physicians to judge how well health AI models handle diagnoses, clinical reasoning and care plans. The work is part-time and asynchronous, aimed at MDs and DOs who currently see patients in primary care, internal medicine or hospital-based specialties.

Posted October 8, 2026

OpenRecently verified$150/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

This short-term remote project asks practicing oncologists to test how frontier AI models reason through real-world cases, with a focus on clinical trial design, patient eligibility and treatment decisions. Contributors write challenging scenarios, judge model answers against current standard-of-care guidelines, and flag errors or unsafe advice. It suits clinicians with fellowship training and hands-on clinical trial experience who want to apply their expertise to AI research.

Posted October 8, 2026

OpenRecently verified$380/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This contract role asks a family medicine physician to annotate and review clinical records and to judge AI-generated medical outputs for accuracy and safety. It suits clinicians who want to help AI systems reason well about patients of all ages, from children to older adults. No prior AI experience is required, since clinical knowledge is the main qualification.

Posted October 8, 2026

OpenRecently verified$100-120/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This contract role asks hospital medicine physicians to review inpatient medical records, including FHIR data and clinical notes, to improve the accuracy of data used to train AI systems. Reviewers confirm that clinical claims are supported by the chart, flag missing or inconsistent details, and apply the same written guidelines throughout. It suits clinicians who know inpatient workflows and documentation standards.

Posted October 8, 2026

OpenRecently verified$100-120/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This contract role asks registered nurses to review inpatient clinical documentation and judge whether AI-generated clinical outputs are accurate, safe and appropriate. No prior AI experience is needed, since bedside nursing knowledge is the main requirement. It suits nurses with broad inpatient clinical exposure who can communicate clearly with project stakeholders.

Posted October 8, 2026

OpenRecently verified$60-70/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Data & Machine Learning
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This part-time contract involves screening business datasets for personal information, checking the accuracy of analytical results, and judging how AI systems answer data analysis questions. It is aimed at experienced data analysts who take quality assurance and precision seriously.

Posted October 8, 2026

OpenRecently verified$50-80/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Business, Consulting & Operations
  • Platform: Ethos
  • Level: Intermediate
  • Commitment: 5-20 hrs/week

You help a leading AI lab train its language model on professional presentation work in healthcare. The role is for strategy consultants with a background in healthcare or life sciences who build executive slide decks and can judge the ones a model produces.

OpenRecently verified$100/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Ethos
  • Level: Expert
  • Commitment: 5-20 hrs/week

You help a leading AI lab train its language model on professional presentation work in program management. The role is for program managers and PMO leaders who own executive reporting and can judge the decks a model produces.

OpenRecently verified$70/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: Canada, United States
  • Level: Expert

This contract asks you to test how well AI-powered enterprise search and workplace assistants answer HR questions, scoring the quality of each answer and the sources behind it. It is aimed at experienced HR practitioners currently working in HR, People Ops, Payroll or HR Systems roles at technology companies.

Posted October 7, 2026

OpenRecently verified$70-90/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This contract role asks hospitalist physicians to review inpatient medical records, including FHIR data and clinical notes, to check that clinical statements are supported by the chart. The work feeds AI training and data quality, using consistent guidelines and clear reporting to the project team. It is aimed at physicians with inpatient medicine backgrounds who can judge documentation accuracy.

Posted October 7, 2026

OpenRecently verified$70-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Data & Machine Learning
  • Platform: micro1
  • Location: 58 countries

This is a remote contract role in which you review and correct bounding boxes that an AI model has produced on still drone and maritime sensor images, and you add labels for objects the model missed. It suits people with solid domain knowledge in aerial or maritime imagery who can follow detailed labeling guidelines, and it requires no prior AI experience.

Posted October 7, 2026

OpenRecently verified$20-60/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Mercor is seeking experienced Administrative Services Managers to evaluate and improve cutting-edge AI systems. You will analyze content and administrative workflows generated by AI, drawing on your operational expertise in planning, coordination, and management of organizational support services such as facilities, document management, procurement, and office operations. This role involves qualitative evaluation of administrative artifacts and improvement recommendations based on your business experience.

Posted October 6, 2026

OpenRecently verified$70-110/hr

Referral link: we may earn a fee. Apply without it

  • Software Engineering
  • Platform: Mercor
  • Location: United States
  • Level: Expert

This is a paid evaluation role for database engineers who will design realistic tasks that test how well AI agents handle database work. Each task is a self-contained problem with a reproducible environment, a working reference solution, tests, and grading criteria that separate robust engineering from code that only looks correct. It is aimed at experts with deep database-engine internals experience, not at DBA, SQL analytics, or ETL-only profiles.

Posted October 6, 2026

OpenRecently verified$2200/task

Referral link: we may earn a fee. Apply without it

  • Legal
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Mercor is recruiting experienced private investigators to evaluate and improve advanced AI systems. You will examine content and workflows generated by AI in the investigations domain, drawing on your real expertise in gathering, analyzing, and reporting facts related to legal, financial, and personal matters. This role is for professionals with a minimum of four years of investigations experience.

Posted October 6, 2026

OpenRecently verified$70-110/hr

Referral link: we may earn a fee. Apply without it

  • Writing, Creative & Design
  • Platform: micro1
  • Location: 58 countries

This job involves evaluating and selecting image collections based on their artistic quality, composition, lighting and originality. You will apply your visual expertise to contribute to the training of next-generation AI models by providing detailed critical feedback. This role is aimed at professionals with solid experience in visual and creative fields.

Posted October 6, 2026

OpenRecently verified$35-105/hr

Referral link to micro1's job list: search for this role there. Open this exact job

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

682 AI Evaluation AI training jobs are open on SideHustler today. 29 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 9, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 640 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (142), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.