AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

682
open jobs
29
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 640 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 9, 2026.

381 to 400 of 682 jobs
  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: Filipino

This job invites a psychologist with a doctoral degree to contribute to training AI systems by leveraging your clinical and academic expertise. You will analyze clinical cases, create culturally adapted psychological scenarios, and evaluate AI-generated content, working bilingually in Tagalog-English to ensure data quality and relevance.

Posted September 29, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: Japanese

You are a bilingual Japanese psychiatrist and contribute to the training of AI systems by developing complex clinical case studies and annotating psychological evaluations. Your professional expertise in Japanese psychiatry and your English proficiency will enable the creation of nuanced, culturally appropriate training data for AI models in the mental health field.

Posted September 28, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: Japanese

This role consists of contributing to the training of AI systems by leveraging your psychological expertise. You will analyze clinical cases, develop culturally adapted psychological scenarios in Japanese and English, and evaluate the ethical and scientific quality of AI-generated content, working remotely in a multidisciplinary environment.

Posted September 28, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Data & Machine Learning
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role involves evaluating and monitoring the quality of datasets intended to train next-generation AI systems. You will examine detailed outputs to verify their accuracy, completeness and compliance with guidelines, applying structured evaluation grids and documenting your analyses. The ideal profile masters data review and possesses solid domain expertise, though AI experience is not required.

Posted September 27, 2026

OpenRecently verified$30-60/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Business, Consulting & Operations
  • Platform: Terac
  • Location: United States

This paid study asks operations professionals who manage internal data at architecture firms to take part in a remote video interview. Participants explain how they turn messy operational data into structured formats used for post-training evaluation, and they review hypothetical data transformation scenarios. It is aimed at data operations leads, BIM managers, and workflow specialists with direct hands-on experience.

Posted September 27, 2026

  • Engineering
  • Platform: Mercor
  • Location: United States
  • Level: Expert

Mercor is seeking experienced electrical and hardware engineers to support an advanced research initiative in engineering and AI. The role consists of evaluating hardware designs, analyzing complex technical decisions, and providing feedback based on deep practical experience in hardware development in production environments.

Posted September 26, 2026

OpenRecently verified$100-120/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

You will work on research physics problems in optics, atomic physics or molecular physics, by creating, solving, reviewing or auditing challenges designed to evaluate the reasoning capabilities of frontier AI models. This role is for physicists who have published on specific phenomena in one of seven areas of specialization and are capable of producing research-quality work with detailed written reasoning.

Posted September 25, 2026

OpenRecently verified$80-110/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

This role involves working on fundamental physics problems at research level for the CritPt benchmark, which evaluates the reasoning capabilities of AI models in advanced theoretical physics. Based on your profile and publications, you will create problems, solve them, review work or verify it in your area of specialization. This role is intended for high energy physicists who have published in one of seven specified domains and who master the relevant methods.

Posted September 25, 2026

OpenRecently verified$80-110/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

This role invites mathematical physicists to contribute to a research benchmark in fundamental physics, designed to assess whether cutting-edge AI models can solve real research problems in physics. You will work on specialized challenges in one of four proposed domains (special functions, integrable systems, permutation combinatorics, conformal geometry), creating, solving, reviewing, or validating problems at the level of academic research.

Posted September 25, 2026

OpenRecently verified$80-110/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

This role involves contributing to the creation and evaluation of research-level quantum physics problems for a benchmark designed to test the reasoning capabilities of state-of-the-art AI models. Depending on your area of specialization and publication record, you may create problems, solve them, review the work of others, or audit it. This work requires solid training in quantum information physics and verifiable publication in one of the twelve specified research domains.

Posted September 25, 2026

OpenRecently verified$80-110/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

This role involves contributing to the creation and evaluation of research-level physics challenges for a benchmark designed to test the physical reasoning capabilities of AI models. The work involves creating problems, solving them, examining solutions or auditing them in your area of specialization in statistical physics. This role is intended for PhDs in statistical physics with published experience on specific phenomena and who master the corresponding mathematical and computational methods.

Posted September 25, 2026

OpenRecently verified$80-110/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

This role involves contributing to a research benchmark in physics aimed at evaluating the reasoning capabilities of AI models on authentic research problems. Depending on your area of specialization and profile, you will develop problems, solve them, review the work of other researchers or audit it, with an emphasis on scientific rigor and clarity of writing.

Posted September 25, 2026

OpenRecently verified$80-110/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: micro1
  • Location: 58 countries

You will evaluate the capabilities of AI assistants in managing daily digital tasks such as email management, calendars, and online bookings. The job consists of testing these tools, verifying the quality of their results, and providing structured feedback to improve AI models. This role is for experienced users of Google Workspace tools looking to contribute to the training of next-generation AI assistants.

Posted September 25, 2026

OpenRecently verified$15-30/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This remote freelance role involves labeling and reviewing data to help improve artificial intelligence models. It suits people with prior experience in data annotation, data labeling, or AI model evaluation who can work independently on a flexible schedule. Tasks range from categorizing text and tagging entities to judging the quality and safety of AI-generated responses.

Posted September 25, 2026

  • Finance & Accounting
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Mercor is recruiting experienced professionals in securities, commodities, and financial services sales to evaluate and improve cutting-edge AI systems. You will analyze content and workflows generated by AI in your field of expertise, drawing on your practical experience with financial markets and client relationships.

Posted September 24, 2026

OpenRecently verified$70-110/hr

Referral link: we may earn a fee. Apply without it

  • Legal
  • Platform: micro1
  • Location: 58 countries
  • Level: Entry level

This role is aimed at junior-level lawyers wishing to participate in the training and evaluation of AI models specialized in contract analysis. You will examine contracts, provide feedback on AI responses, and contribute to improving the accuracy of these systems in interpreting legal language and identifying risks. The role combines traditional transactional legal practice with applied AI research in the legal field.

Posted September 24, 2026

OpenRecently verified$70-90/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: Mercor
  • Location: United States

Mercor is recruiting operational cybersecurity experts to participate in video research interviews designed to create a benchmark of AI agents in enterprise defense. This short, well-paid job involves sharing your practical experience on tools, decision-making processes, and the potential impact of AI in your field.

Posted September 23, 2026

OpenRecently verified$125-175/hr

Referral link: we may earn a fee. Apply without it

  • Legal
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

You are an experienced technology law attorney joining a pioneering team that trains AI models to master contract analysis and negotiation. You participate in contract simulation exercises and evaluate AI system responses to improve their ability to identify legal risks and interpret clauses with precision.

Posted September 23, 2026

OpenRecently verified$85-105/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: micro1
  • Location: 58 countries

This role involves leveraging your cybersecurity expertise to train next-generation AI systems. You will use Security Onion to execute security workflows, analyze alerts and network traffic, then document your approaches to create quality training data for AI models. The ideal profile combines solid hands-on experience with Security Onion and the ability to clearly communicate technical steps.

Posted September 23, 2026

OpenRecently verified$90-175/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries

This role is aimed at experienced chartered accountants tasked with training AI models by leveraging your business expertise. You will annotate and evaluate financial documents, tax returns, and audit procedures to enable AI systems to learn from real-world cases. No prior knowledge of artificial intelligence is required, only your accounting specialization matters.

Posted September 23, 2026

OpenRecently verified$60-120/hr

Referral link to micro1's job list: search for this role there. Open this exact job

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

682 AI Evaluation AI training jobs are open on SideHustler today. 29 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 9, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 640 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (142), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.