Clinical Evaluation AI training jobs

Open AI training and model evaluation roles that call for Clinical Evaluation expertise, gathered from 7 platforms. Part of Health & Medicine.

58
open jobs
7
posted this week
$85/hr
median published rate
$380/hr
highest published rate
7
platforms hiring

Pay figures cover the 47 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 8, 2026.

1 to 20 of 58 jobs
  • Health & Medicine
  • Platform: Alignerr

This remote, hourly contract asks you to review and grade AI-generated content about healthcare data, electronic health record systems, and clinical workflows. The work helps AI labs make their models more accurate on medical information. It suits health informatics professionals who can judge clinical and technical accuracy and explain their feedback in writing.

  • Health & Medicine
  • Platform: Alignerr

This remote contract role asks public health graduates to bring their domain knowledge to AI training. The work involves building realistic population health scenarios, writing evidence-based reference answers, and reviewing AI-generated health advice for accuracy, logic, bias, and ethics. It suits independent, self-directed specialists who can explain health data clearly, and no prior AI experience is required.

  • Health & Medicine
  • Platform: Alignerr

This remote contract role asks medical science professionals to review and validate AI-generated biomedical content for accuracy, clarity, and context. It suits people with a background in medical affairs, clinical research, or scientific communication who want to shape how AI handles medical information.

  • Health & Medicine
  • Platform: Alignerr

This remote freelance role asks registered nurses and clinical informatics professionals to review AI-generated clinical content, EHR workflows and health information outputs. The work is to judge their accuracy, safety and fit with real nursing practice, then give structured feedback that helps healthcare AI systems behave more reliably. It suits clinicians with hands-on health IT experience who can work independently on task-based assignments.

  • Health & Medicine
  • Platform: Alignerr

This remote contract role asks oncology clinical research experts to review AI-generated cancer content for scientific accuracy, clinical validity, and regulatory alignment. The work covers trial protocols, safety and efficacy data, biomarker findings, and FDA or EMA-style reporting. It is aimed at professionals with hands-on experience designing and analyzing oncology trials who want to work asynchronously on their own schedule.

  • Health & Medicine
  • Platform: Alignerr

This is a remote, hourly contract for senior clinical scientists who will design and check clinical trial protocols and audit trial results. The work also includes judging whether AI-generated clinical analyses are scientifically sound and giving written feedback so that AI models reason more accurately about clinical evidence. It suits experienced professionals who have worked on regulatory submissions and can work independently on their own schedule.

  • Health & Medicine
  • Platform: Alignerr

This remote, hourly contract involves labeling and reviewing clinical data, including electronic health records, clinical notes, lab results and medical codes, to help train AI systems used in healthcare. Work includes quality assurance checks, flagging ambiguous cases and giving expert feedback on model outputs. It is aimed at clinicians, health informaticists and healthcare data professionals who can work independently on a flexible schedule.

  • Health & Medicine
  • Platform: DataAnnotation

This role involves evaluating how well AI models reason through cardiology cases, identifying errors in clinical judgment that could harm patients. Cardiologists will write prompts to test AI reasoning, review AI outputs for accuracy against current guidelines, and provide correct answers when the model errs, serving as the training signal for AI improvement.

Talent poolRecently verified$40-125/hr
  • Health & Medicine
  • Platform: DataAnnotation

This role involves evaluating AI-generated responses to health and medical questions for clinical accuracy and safety. The expert identifies errors in reasoning, missed contraindications, unsafe advice, and writes correct answers grounded in current clinical practice, ensuring AI models provide guidance that real clinicians would trust.

Talent poolRecently verified$40-125/hr
  • Health & Medicine
  • Platform: DataAnnotation

This role involves evaluating how AI models handle nursing judgment in clinical scenarios such as patient assessment, medication safety, and triage protocols. Nurses will identify unsafe reasoning, incorrect clinical priorities, and protocol violations, then provide corrections grounded in real bedside practice.

Talent poolRecently verified$40-125/hr
  • Health & Medicine
  • Platform: DataAnnotation

A physician evaluates AI model responses on clinical questions against current standards of care, identifying unsafe or outdated guidance and writing correct clinical answers when needed. This role requires assessing AI clinical reasoning and creating prompts that test differential diagnosis, treatment planning, and patient communication.

Talent poolRecently verified$40-125/hr
  • Health & Medicine
  • Platform: DataAnnotation

This role asks radiologists and imaging technologists to read de-identified scans and judge how accurately frontier AI models interpret them. Contributors write structured reports and impressions, then compare them with model output to find errors and gaps. It suits clinicians with hands-on imaging experience who can explain why a machine-generated report is wrong.

Talent poolRecently verified$40-125/hr
  • Health & Medicine
  • Platform: DataAnnotation

This role involves evaluating AI models' clinical reasoning on nursing topics including assessment, medication administration, and patient care protocols. Registered nurses will identify gaps between textbook knowledge and real bedside practice, write test prompts to probe nursing judgment, and provide corrected guidance when AI outputs contain unsafe or incorrect clinical information.

Talent poolRecently verified$40-125/hr
  • Health & Medicine
  • Platform: Mercor
  • Location: United States

This role asks practicing physicians to judge how well health AI models handle diagnoses, clinical reasoning and care plans. The work is part-time and asynchronous, aimed at MDs and DOs who currently see patients in primary care, internal medicine or hospital-based specialties.

Posted October 8, 2026

OpenRecently verified$150/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

This short-term remote project asks practicing oncologists to test how frontier AI models reason through real-world cases, with a focus on clinical trial design, patient eligibility and treatment decisions. Contributors write challenging scenarios, judge model answers against current standard-of-care guidelines, and flag errors or unsafe advice. It suits clinicians with fellowship training and hands-on clinical trial experience who want to apply their expertise to AI research.

Posted October 8, 2026

OpenRecently verified$380/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This contract role asks hospitalist physicians to review inpatient medical records, including FHIR data and clinical notes, to check that clinical statements are supported by the chart. The work feeds AI training and data quality, using consistent guidelines and clear reporting to the project team. It is aimed at physicians with inpatient medicine backgrounds who can judge documentation accuracy.

Posted October 7, 2026

OpenRecently verified$70-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This part-time contract asks physicians to review clinical cases, patient data, and AI-generated medical content, giving expert feedback that helps train AI systems supporting medical decisions. It is aimed at practicing doctors, especially those from general specialties such as internal medicine, family practice, or emergency medicine. No prior AI experience is needed, since clinical knowledge is the main requirement.

Posted October 6, 2026

OpenRecently verifiedPay not disclosed

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: Mindrift
  • Location: 36 countries
  • Level: Intermediate

A US-based licensed physician contributes to AI training projects by evaluating and annotating AI-generated medical content, creating complex clinical scenarios, and assessing the accuracy and safety of AI responses. This is part-time, project-based work improving healthcare AI systems.

Posted October 5, 2026

OpenRecently verifiedPay not disclosed
  • Health & Medicine
  • Platform: Mindrift
  • Location: United States
  • Level: Intermediate

This is a project-based AI training role for practicing US physicians and licensed clinical professionals to evaluate and improve health-focused AI systems. Contributors review AI-generated medical content for accuracy and safety, create complex clinical scenarios to test AI reasoning, and develop both detailed and ambiguous versions of medical tasks to train AI agents on realistic clinical complexity.

Posted October 5, 2026

OpenRecently verifiedPay not disclosed
  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This role calls for a bilingual Cantonese psychiatrist to contribute to the training of next-generation AI systems. You will put your clinical expertise to work creating culturally nuanced psychiatric scenarios, producing detailed evaluations and annotations in Cantonese and English. The project requires a solid understanding of Cantonese psychiatric practices and the ability to apply clinical ethics when reviewing sensitive content.

Posted October 3, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about Clinical Evaluation AI training jobs

How many Clinical Evaluation AI training jobs are open right now?

58 Clinical Evaluation AI training jobs are open on SideHustler today. 7 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 8, 2026.

How much do Clinical Evaluation AI training jobs pay?

Among the 47 open roles that publish an hourly rate in USD, pay runs from $30 to $380 per hour, with a median of $85. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer Clinical Evaluation AI training jobs?

micro1 (20), Mercor (17), Alignerr (7), DataAnnotation (6), Vetto (5), Mindrift (2) and Terac (1). You apply on the platform itself, which handles screening, contracts and payment.

What skills do Clinical Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are Clinical Evaluation, AI Evaluation, Psychiatry & Psychology, Clinical Documentation, Clinical Research. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.