AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

682
open jobs
29
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 640 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 9, 2026.

341 to 360 of 682 jobs
  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: Russian

You are a bilingual psychiatrist who will contribute to training AI systems by developing and annotating complex psychiatric case studies that reflect Russian clinical practices and perspectives. This role will allow you to share your medical expertise to shape AI models' understanding in the mental health field.

Posted October 2, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: Russian

This role requires a psychologist with a doctoral degree to bring clinical and academic expertise to refine AI model learning. You will analyze psychological cases, develop bilingual Russian-English pedagogical scenarios, and assess the ethical and scientific quality of content generated by AI systems, leveraging your nuanced understanding of both cultures and languages.

Posted October 2, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: Vietnamese

This role calls for bilingual Vietnamese psychiatrists to contribute to the training of AI systems. You will develop complex psychiatric case studies and produce clinical annotations in Vietnamese and English, integrating Vietnamese psychiatric perspectives and practices. Your expertise in mental health will enable AI models to better understand psychiatric concepts with cultural nuance.

Posted October 2, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: Vietnamese

This role invites a psychologist with a PhD and fluency in Vietnamese and English to contribute to training AI models. You will analyze psychological case studies, develop culturally adapted scenarios in both languages, and evaluate the quality and ethics of content generated by AI systems. No prior experience in artificial intelligence is necessary.

Posted October 2, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Engineering
  • Platform: micro1
  • Location: 58 countries
  • Level: Expert

This role invites you to share your CAD expertise to train next-generation AI systems. You will analyze product requirements, design sophisticated multi-part assemblies with advanced parametric tools, and document your design intentions to guide model learning. You are a CAD expert with at least five years of professional experience mastering major parametric software.

Posted October 2, 2026

OpenRecently verified$40-90/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries

This role invites investment banking professionals to share their expertise to train next-generation AI systems. You will document how AI tools integrate into your daily workflows of financial modeling, research and transaction analysis, providing critical feedback on their effectiveness and limitations.

Posted October 2, 2026

OpenRecently verified$150-220/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Engineering
  • Platform: micro1
  • Location: 58 countries
  • Level: Expert

This role involves helping train AI systems by analyzing complex mechanical assemblies and formulating demanding technical questions that test reasoning in mechanical design. You will validate answers directly from CAD models, examine AI-generated results to identify weaknesses, and document your findings rigorously. This position is suited to senior mechanical engineers with proven expertise in multi-part assembly design and mastery of parametric CAD tools.

Posted October 2, 2026

OpenRecently verified$40-90/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: Terac
  • Location: United States

This remote study asks practicing software engineers to review programming tasks and the evaluation harnesses built to test AI agents on them. You will check whether the tasks, test cases and environment structure reflect realistic software engineering work. It suits engineers with solid experience building, testing and reviewing complex systems.

Posted October 2, 2026

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: India
  • Level: Entry level

This role involves evaluating AI assistants live during India-involved T20I cricket matches by asking fan questions and rating the quality, accuracy, and presentation of their responses. The work takes place only on scheduled match days, in two to two and a half hour slots, and requires solid knowledge of international cricket.

Posted October 1, 2026

OpenRecently verified$20/task

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Board-certified emergency medicine physician sought to evaluate and annotate clinical documentation generated by medical AI tools. The role involves reviewing patient records, validating the accuracy of emergency summaries produced by AI, and flagging errors or omissions to help engineering teams improve these systems. Profile: practicing emergency physician with at least 2 to 3 years of experience.

Posted October 1, 2026

OpenRecently verified$150-170/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States
  • Language: Spanish

You will evaluate AI assistants live during football matches, asking natural fan questions and rating the quality, accuracy, and clarity of their responses. This role requires genuine passion for football and availability during scheduled match times.

Posted October 1, 2026

OpenRecently verified$105-140/hr

Referral link: we may earn a fee. Apply without it

  • Legal
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role is designed for lawyers specializing in derivatives to contribute to the training and evaluation of AI models applied to financial law. You will participate in negotiation and review exercises on ISDA contracts, then provide expert feedback to improve the performance of AI systems in interpreting derivatives documentation. Your work will help develop more accurate and reliable legal tools.

Posted October 1, 2026

OpenRecently verified$150-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role asks a Canadian personal finance expert to train AI models by performing authentic domestic financial management tasks and evaluating advice generated by these systems. You will apply your knowledge of Canadian savings products and tax matters to produce quality feedback intended to improve the financial reasoning of the models.

Posted October 1, 2026

OpenRecently verified$90-110/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

You bring your expertise in personal finance to improve AI systems by carrying out concrete domestic financial management tasks in the British context. This role is suited to experienced professionals in banking, wealth management, accounting or tax advisory, able to assess and improve the quality of advice generated by learning models.

Posted October 1, 2026

OpenRecently verified$90-110/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: Terac
  • Location: Argentina

This paid study asks experienced software engineers to review proposed programming tasks and the harnesses that test AI agents, checking their logic, structure and realism. The work takes place in a screen-shared session that mixes code review, technical discussion and direct feedback on task design. It is aimed at engineers who have already built, reviewed or tested evaluation harnesses or coding tasks.

Posted October 1, 2026

  • Software Engineering
  • Platform: Terac
  • Location: India

This is a paid research interview in which experienced software engineers review proposed programming tasks and their evaluation environments. Participants assess how realistic, difficult and well structured the challenges are, and they suggest ways to make the test suites more robust. It suits engineers who already build or review complex codebases and know how test harnesses work.

Posted October 1, 2026

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

Mercor is looking for clinical mental health experts to evaluate the realism and quality of simulated clinical conversations generated by AI. This short-term role, entirely remote, is aimed at mental health professionals capable of applying their clinical experience to judge the authenticity of realistic conversational scenarios.

Posted September 30, 2026

OpenRecently verified$100/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States

You will contribute to training AI models by writing and validating formalized mathematical proofs in Lean 4. You convert informal mathematical statements into precise declarations and verify that generated proofs are correct both formally and mathematically. This role is aimed at experienced Lean engineers or mathematicians with strong expertise in formal proof.

Posted September 30, 2026

OpenRecently verified$90-110/hr

Referral link: we may earn a fee. Apply without it

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries

This job consists of evaluating and annotating U.S. tax return processes and financial workflows to train AI systems. You must apply your tax expertise to assess the accuracy and completeness of these workflows, comparing system-generated results with real documents and scenarios.

Posted September 30, 2026

OpenRecently verified$30-110/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This job offers you the opportunity as a pharmaceutical expert to evaluate and create high-quality pharmaceutical content and data for a technology project. You will leverage your professional expertise to examine the scientific and clinical accuracy of information related to medications, clinical research, and product safety.

Posted September 30, 2026

OpenRecently verified$80-285/hr

Referral link to micro1's job list: search for this role there. Open this exact job

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

682 AI Evaluation AI training jobs are open on SideHustler today. 29 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 9, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 640 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (142), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.