AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

679
open jobs
26
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 637 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 8, 2026.

181 to 200 of 679 jobs
  • Science & Research
  • Platform: Alignerr

This contract role asks a physics specialist to label and organize physics material used to train AI models. The work includes reviewing AI-generated solutions to physics problems and helping build training datasets across topics and difficulty levels. It suits candidates with research, teaching or applied problem-solving experience in physics.

  • Science & Research
  • Platform: Alignerr

This is a remote, hourly contract role for physicists who want to help AI models reason through hard science problems. The work centers on writing challenging physics problems with worked solutions and on checking how AI systems handle them, from undergraduate to PhD level. No prior AI background is needed, only strong physics knowledge and clear written communication.

  • Science & Research
  • Platform: Alignerr

This is a remote, hourly contract for physics graduates and current Masters students who want to help train and evaluate advanced AI models. The work is writing physics problems, solving them step by step, and judging whether AI-generated answers are scientifically accurate and well reasoned.

  • Science & Research
  • Platform: Alignerr

This remote contract asks a physics expert to write advanced problems that test how well AI models reason, along with clear step-by-step solutions. The work also includes checking AI-generated answers for accuracy and quality of reasoning, and helping researchers refine benchmarks. It is aimed at people holding a master's degree in physics or a related field, and no prior AI experience is needed.

  • Science & Research
  • Platform: Alignerr

This freelance contract role asks physics specialists to design hard problems and write clear step-by-step solutions that help train AI models. The work also includes checking AI-generated answers for scientific accuracy and quality of reasoning. It suits physicists with a master's-level background who can work independently and asynchronously.

  • Health & Medicine
  • Platform: Alignerr

This contract role asks population health experts to review and improve AI systems that interpret public health data. The work involves checking AI-generated trends, dashboards, and recommendations against evidence-based standards. It suits professionals with a background in health data and analytics who can work independently and remotely.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Portuguese

This remote, part-time freelance role asks Portuguese speakers to review text produced by AI systems and content written in Portuguese. The work centers on judging whether the language is accurate, natural, and culturally suited to its audience. It suits people with deep knowledge of Brazilian or European Portuguese who can give structured, independent feedback.

  • Data & Machine Learning
  • Platform: Alignerr

This remote contract role asks you to forecast the likelihood of geopolitical, financial, and industry events and to explain each estimate in writing. The work helps AI research teams build models that reason more accurately about uncertainty. It suits analytically minded people who follow the news closely and think in probabilities.

  • Health & Medicine
  • Platform: Alignerr

This is a remote, hourly contract for senior clinical scientists who will design and check clinical trial protocols and audit trial results. The work also includes judging whether AI-generated clinical analyses are scientifically sound and giving written feedback so that AI models reason more accurately about clinical evidence. It suits experienced professionals who have worked on regulatory submissions and can work independently on their own schedule.

  • Software Engineering
  • Platform: Alignerr
  • Level: Expert

This contract role centers on building and tuning high-performance Rust systems that support AI model development, including data pipelines, annotation tooling and evaluation infrastructure. It is aimed at senior engineers with deep Rust expertise who want to work fully remotely alongside research and engineering teams.

  • Data & Machine Learning
  • Platform: Alignerr

This remote, hourly contract role involves checking AI-written product reviews, comparisons, and recommendations for accuracy, honesty, and usefulness. It suits people who know online shopping well and have a critical eye, and no retail or AI background is needed.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: German

This remote freelance role asks German-speaking voice professionals to record scripted and creative performances and to judge AI-generated German speech for naturalness and pronunciation. The recordings and feedback help train AI voice systems that will speak to German audiences. It suits experienced voice actors who work from a home studio and have a critical ear for delivery.

OpenRecently verified$250-280/hr
  • Writing, Creative & Design
  • Platform: Alignerr

This remote contract role asks you to write original, inventive prompts that test how advanced AI models reason, respond, and handle edge cases. It suits people with strong writing skills and broad curiosity, and no technical or AI background is needed.

  • Science & Research
  • Platform: Alignerr
  • Level: Intermediate

This contract role asks economists to review and grade the economic reasoning in content drafted by AI systems, with a focus on regulation, welfare, taxation and public programs. It suits professionals with a background in public or applied economics who can give precise, written feedback on arguments and policy conclusions while working independently on task-based assignments.

  • Software Engineering
  • Platform: Alignerr
  • Level: Intermediate

This contract role asks Python developers to assess code produced by AI systems for correctness, efficiency and good practice, and to solve algorithmic problems by writing solutions from scratch. It suits senior developers who can explain code logic clearly and work independently on a flexible, remote schedule.

  • Software Engineering
  • Platform: Alignerr
  • Level: Intermediate

This remote contract role asks senior Python developers to judge code written by AI systems for correctness, efficiency and professional standards. The work also includes solving algorithmic problems and explaining the reasoning clearly so that AI models learn to write better Python. It suits experienced engineers who can work independently on flexible hours.

  • Software Engineering
  • Platform: Alignerr
  • Level: Intermediate

This contract role involves building and tuning Python systems that feed data pipelines, annotation tools and model evaluation workflows for AI teams. It suits senior engineers who can work full-stack and backend, with a strong focus on reliability and clean, well-tested code.

  • Data & Machine Learning
  • Platform: Alignerr

This contract role asks quantitative professionals to review and improve the mathematical and statistical outputs that AI models produce, especially financial models, risk estimates and forecasts. The work is remote, hourly and self-scheduled, and suits people with a solid background in quantitative finance, statistics or applied mathematics who notice when a result does not hold up.

  • Finance & Accounting
  • Platform: Alignerr

This remote, hourly contract role involves reviewing the financial models and forecasts that AI systems generate, testing their assumptions and checking their logic. It is aimed at professionals with a solid background in quantitative finance, statistics or data analysis who can work independently on task-based assignments.

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

679 AI Evaluation AI training jobs are open on SideHustler today. 26 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 8, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 637 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (139), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.