AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

682
open jobs
29
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 640 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 9, 2026.

321 to 340 of 682 jobs
  • Business, Consulting & Operations
  • Platform: micro1
  • Location: 58 countries

You will collaborate on a project to train AI assistants using your pedagogical expertise and knowledge of literature and composition. You will simulate realistic teacher tasks, give instructions to an AI assistant, then evaluate its responses on language quality, adherence to directives, and stylistic appropriateness. This role draws on your classroom experience and your ability to analyze texts in detail.

Posted October 6, 2026

OpenRecently verified$15-40/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries

This role invites equity research analysts with recent sell-side experience to contribute to training financial AI models. You will participate in expert interviews, test AI-powered tools, and provide detailed feedback on their relevance and usability for equity research. Your domain expertise in equity research will shape how AI systems learn to analyze financial markets.

Posted October 6, 2026

OpenRecently verified$180-220/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This part-time contract asks physicians to review clinical cases, patient data, and AI-generated medical content, giving expert feedback that helps train AI systems supporting medical decisions. It is aimed at practicing doctors, especially those from general specialties such as internal medicine, family practice, or emergency medicine. No prior AI experience is needed, since clinical knowledge is the main requirement.

Posted October 6, 2026

OpenRecently verifiedPay not disclosed

Referral link to micro1's job list: search for this role there. Open this exact job

  • Legal
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role is for experienced attorneys in technology contract law who want to contribute to the development of legal AI models. You will evaluate AI responses against contractual scenarios, refine evaluation frameworks, and collaborate with cross-functional teams to improve AI-assisted contract review systems.

Posted October 6, 2026

OpenRecently verified$80-110/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: micro1
  • Location: 58 countries

You will contribute to training next-generation AI models by creating realistic coding tasks and implementing deterministic verifiers. This flexible contracting role is aimed at experienced backend engineers capable of generating relevant technical challenges and automatically evaluating solutions.

Posted October 6, 2026

OpenRecently verified$30-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Engineering
  • Platform: Mindrift
  • Location: 61 countries

This role involves training and evaluating AI agents on fluid simulation tasks by setting up CFD simulations, creating reference solutions, and validating agent outputs. It is designed for experienced mechanical or aerospace engineers with strong computational fluid dynamics expertise.

Posted October 6, 2026

OpenRecently verifiedPay not disclosed
  • Engineering
  • Platform: Mindrift
  • Location: United States

This role involves training and evaluating AI agents on expert-level computational fluid dynamics tasks. You will set up and run simulations using industry-standard CFD tools, create verified reference solutions, and develop Python validation checks to assess agent performance on fluid simulation problems.

Posted October 6, 2026

OpenRecently verifiedPay not disclosed
  • Legal
  • Platform: micro1
  • Location: 58 countries

This role is for experienced litigation attorneys holding a US license who wish to contribute to the development of artificial intelligences applied to the legal field. You will evaluate the performance of AI models on litigation tasks, provide expert feedback, and create rigorous evaluation frameworks to improve the accuracy and legal judgment of these systems.

Posted October 5, 2026

OpenRecently verified$80-150/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Engineering
  • Platform: Mindrift
  • Location: 60 countries
  • Level: Expert

This role involves training and evaluating AI agents that perform mechanical engineering tasks within a CAD environment. You will design part geometry, create reference solutions, and assess how well the AI completes complex engineering work that would normally take an expert many hours.

Posted October 5, 2026

OpenRecently verifiedPay not disclosed
  • Health & Medicine
  • Platform: Mindrift
  • Location: 36 countries
  • Level: Intermediate

A US-based licensed physician contributes to AI training projects by evaluating and annotating AI-generated medical content, creating complex clinical scenarios, and assessing the accuracy and safety of AI responses. This is part-time, project-based work improving healthcare AI systems.

Posted October 5, 2026

OpenRecently verifiedPay not disclosed
  • Health & Medicine
  • Platform: Mindrift
  • Location: United States
  • Level: Intermediate

This is a project-based AI training role for practicing US physicians and licensed clinical professionals to evaluate and improve health-focused AI systems. Contributors review AI-generated medical content for accuracy and safety, create complex clinical scenarios to test AI reasoning, and develop both detailed and ambiguous versions of medical tasks to train AI agents on realistic clinical complexity.

Posted October 5, 2026

OpenRecently verifiedPay not disclosed
  • Engineering
  • Platform: Mindrift
  • Location: United States
  • Level: Expert

This freelance role involves training and evaluating AI agents to perform complex CAD design tasks in professional engineering software environments. You will create part geometry, develop reference solutions, and assess AI agent performance on projects that would typically require 10-15 hours of expert work.

Posted October 5, 2026

OpenRecently verifiedPay not disclosed
  • Writing, Creative & Design
  • Platform: Mercor
  • Location: United States

You will compare images generated by artificial intelligence based on the same prompt and designate the most successful image. This qualitative evaluation role requires careful analysis according to several criteria (faithfulness to the prompt, visual quality, anatomy, rendered text, artifacts) and the writing of a concise justification for your choice.

Posted October 3, 2026

OpenRecently verified$30/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This role calls for a bilingual Cantonese psychiatrist to contribute to the training of next-generation AI systems. You will put your clinical expertise to work creating culturally nuanced psychiatric scenarios, producing detailed evaluations and annotations in Cantonese and English. The project requires a solid understanding of Cantonese psychiatric practices and the ability to apply clinical ethics when reviewing sensitive content.

Posted October 3, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: Chinese

This role invites a psychologist with a doctorate to contribute to training next-generation AI systems. You will analyze psychological case studies, develop culture-sensitive scenarios in Cantonese and English, and evaluate AI-generated content according to ethical and professional standards. No prior AI experience is required - only your expertise in psychology matters.

Posted October 3, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: micro1
  • Location: 58 countries

You will help train AI models by creating realistic programming tasks and deterministic validators. This flexible contractor role will allow you to leverage your backend development experience to design relevant challenges for AI to solve.

Posted October 3, 2026

OpenRecently verified$30-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

This role involves evaluating and improving how AI models respond to sensitive conversations about relationships, emotional well-being, and personal beliefs. You will analyze whether the model's responses remain balanced, neutral, and safe, and you will help researchers define quality standards. This role is suited for professionals from mental health, counseling, or social work who can clearly articulate why one response is better than another.

Posted October 2, 2026

OpenRecently verified$45-70/hr

Referral link: we may earn a fee. Apply without it

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

682 AI Evaluation AI training jobs are open on SideHustler today. 29 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 9, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 640 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (142), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.