AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

679
open jobs
26
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 637 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 8, 2026.

501 to 520 of 679 jobs
  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: Japanese, French, Filipino, Malay, Lithuanian

This role invites developmental psychologists to contribute to the training of AI systems by applying your academic and clinical expertise. You will evaluate and annotate data related to child development, provide critical analyses on their psychological and cultural relevance, and collaborate with interdisciplinary teams to ensure the psychological integrity of AI models.

Posted September 10, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Engineering
  • Platform: micro1
  • Location: 58 countries

This role involves operating a portable robotic manipulator arm in a laboratory to perform repetitive manipulation tasks for training robotic artificial intelligence models. You will execute precise instructions, reset environments between attempts, and report any technical issues to the project lead. No prior artificial intelligence experience is required.

Posted September 10, 2026

OpenRecently verified$30/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Business, Consulting & Operations
  • Platform: micro1
  • Location: 58 countries
  • Level: Expert

You leverage your deep expertise in Shopify to evaluate and improve training scenarios for AI systems. You examine e-commerce task structures, identify gaps with real workflows, and provide detailed feedback to ensure training content is realistic and complete.

Posted September 10, 2026

OpenRecently verified$60-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: Terac
  • Location: United States

This paid study asks practicing dermatologists to judge how well AI diagnostic tools perform on skin images and clinical records. Participants review de-identified cases and check the automated findings against established reference standards. It suits clinicians who routinely diagnose skin conditions and can assess how reliable generated diagnoses are.

Posted September 10, 2026

  • Science & Research
  • Platform: Mercor
  • Location: United States

You are a chemistry expert with practical experience in synthesis, analytical chemistry, or chemical safety. You participate in evaluating AI models by writing calibrated test prompts across three risk levels, then assessing model responses against an established safety framework. This role requires both advanced technical expertise and the ability to document your judgments in an accessible way.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

This role involves evaluating an AI model's responses to requests involving energetic materials and explosives, distinguishing between legitimate questions and dangerous requests. You will write progressively more complex test prompts, assess the model's responses against a defined safety policy, and provide written justification for your technical judgments. This role is intended for experts in materials chemistry specializing in propulsion and initiation systems, with solid practical experience in the field.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

Mercor is looking for nuclear engineering and nuclear safeguards experts to test the capabilities of frontier AI models to assess diversion risk. You will write complex prompts in your field, evaluate responses against a defined policy, and provide written technical justifications. This role requires deep expertise in fuel cycle or non-proliferation, and demonstrated ability to communicate technical issues with clarity.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

Mercor is recruiting radiological safety experts to test the robustness of AI models against dual-use requests. You will write progressively complex prompts, evaluate model responses according to a defined policy, and document reference responses with technical justifications. This role requires deep expertise capable of distinguishing legitimate professional questions from potentially dangerous requests.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: micro1
  • Location: 58 countries

This role invites experienced data analysts to examine and validate large business datasets to contribute to the training of AI systems. The job combines critical data analysis, strict confidentiality compliance, and formulating recommendations to ensure the quality of information for machine learning.

Posted September 9, 2026

OpenRecently verified$31-60/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Writing, Creative & Design
  • Platform: micro1
  • Location: 58 countries

You will contribute to training AI models by evaluating and selecting images according to rigorous aesthetic criteria. Your visual expertise will enable you to judge the quality of composition, lighting, color and originality, while documenting your analyses to calibrate quality standards with other specialists.

Posted September 8, 2026

OpenRecently verified$21-70/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: micro1
  • Location: 58 countries

This role consists of creating reinforcement learning environments to test the capabilities of AI models in solving complex software engineering problems, including code correction, feature creation, and performance optimization. As an expert software engineer, you will produce quality code examples, debugging strategies, and reproducible reference solutions. The role requires confirmed expertise in programming and an established presence in open source.

Posted September 7, 2026

OpenRecently verified$50-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: micro1
  • Location: 58 countries

Create training environments and test cases to evaluate the capabilities of AI models in solving software engineering problems such as code correction, feature development, or performance optimization. This role is aimed at experienced software engineers with recognized expertise in at least one modern programming language.

Posted September 7, 2026

OpenRecently verified$50-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Legal
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role invites legal specialists in contract law to contribute to a confidential project combining law and artificial intelligence. You will apply your transactional expertise to training next-generation AI systems by performing review and drafting tasks on commercial contracts, without needing prior knowledge of AI.

Posted September 6, 2026

OpenRecently verified$90-110/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States
  • Language: Croatian

This role involves evaluating and improving the safety of AI models by analyzing their behavior on sensitive topics in Croatian. You will write expert prompts, classify content according to structured guidelines, and identify adversarial formulations, as an English-Croatian bilingual speaker with no prior AI experience required.

Posted September 4, 2026

OpenRecently verified$38-42/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States
  • Language: Indonesian

This remote-style role asks PhD scientists who are fluent in both English and Indonesian to help make AI models safer on technical subjects. Contributors write expert prompts in Indonesian, then judge model answers for accuracy, helpfulness, and handling of sensitive material. No prior AI experience is needed because the workflow is taught on the job.

Posted September 4, 2026

OpenRecently verified$19-23/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Turkish

You will participate in improving AI model safety by evaluating how they handle sensitive subjects in Turkish. Your linguistic and cultural judgments will help identify and strengthen weaknesses in these systems when facing delicate content. No prior AI experience is required.

Posted September 4, 2026

OpenRecently verified$23-27/hr

Referral link: we may earn a fee. Apply without it

  • Generalist & Data Labeling
  • Platform: Mercor
  • Location: United States
  • Language: Ukrainian

This role involves contributing to AI model safety as a bilingual English-Ukrainian expert. You will evaluate how these models handle sensitive topics in Ukrainian, by writing expert prompts and classifying conversations according to structured guidelines. No prior experience in AI or machine learning is required.

Posted September 4, 2026

OpenRecently verified$38-42/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: Mercor
  • Location: United States
  • Language: Portuguese

You will be hired as a native speaker of Brazilian Portuguese to interact by voice with a next-generation AI assistant and evaluate its responses. You will compose messages, emails and notes via common applications, then rate the quality of the AI interactions in your language.

Posted September 4, 2026

OpenRecently verified$12.6/task

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role involves improving operational efficiency and data quality within complex SaaS environments by documenting and optimizing revenue management processes. The ideal candidate brings deep hands-on experience with CRM systems and adjacent SaaS tools, particularly in regulated contexts, and will contribute to training next-generation AI systems through high-quality business insights.

Posted September 4, 2026

OpenRecently verified$50-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

This role involves evaluating and annotating clinical documents as well as notes generated by medical AI systems, applying your expertise in coding and clinical documentation. You will participate in a multilingual team responsible for ensuring the accuracy and reliability of AI tools in healthcare settings. The ideal candidate must possess recognized medical certification and be proficient in English as well as at least one other language at an advanced level.

Posted September 3, 2026

OpenRecently verified$45-65/hr

Referral link: we may earn a fee. Apply without it

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

679 AI Evaluation AI training jobs are open on SideHustler today. 26 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 8, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 637 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (139), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.