AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

679
open jobs
26
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 637 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 8, 2026.

581 to 600 of 679 jobs
  • Languages, Translation & Voice
  • Platform: micro1
  • Location: 58 countries
  • Language: Italian

You are an Italian bilingual expert and participate in an AI model training project by evaluating the quality and authenticity of statements in Italian. You listen to audio excerpts, analyze fluency and pronunciation, then document your detailed observations in English to improve AI systems.

Posted August 13, 2026

OpenRecently verified$30-65/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Languages, Translation & Voice
  • Platform: micro1
  • Location: 58 countries
  • Language: Russian

This role involves evaluating audio recordings in Russian for an AI training project. You will listen to voice clips, analyze linguistic fluency and authenticity, and provide detailed feedback in English. This position is suited to a native Russian expert with knowledge in linguistics, phonetics, or related fields.

Posted August 13, 2026

OpenRecently verified$30-65/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Languages, Translation & Voice
  • Platform: micro1
  • Location: 58 countries
  • Language: Turkish

You will evaluate the quality and linguistic authenticity of audio recordings in Turkish, with the aim of improving the training of artificial intelligence models. Your analyses will guide the learning of AI systems on understanding and generating the Turkish language. No prior experience in AI is necessary, but native proficiency in Turkish is essential.

Posted August 13, 2026

OpenRecently verified$30-65/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States

You have deployed and operated language model-based agents in production, with direct experience of their failures and real adoption in enterprise environments. This role consists of sharing your expertise on reliability, evaluation, and scaling of AI agents in complex environments, through conversational interviews drawing on your concrete practical feedback.

Posted August 12, 2026

OpenRecently verified$100-500/task

Referral link: we may earn a fee. Apply without it

  • Writing, Creative & Design
  • Platform: Mercor
  • Location: United States
  • Language: Norwegian

You are invited to evaluate the quality of music and lyrics generated by AI models, comparing these creations with detailed standards and published works. This role requires confirmed musical expertise and mastery of Norwegian to analyze the authenticity, creativity, and alignment of productions with the provided prompts.

Posted August 12, 2026

OpenRecently verified$42-78/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: Mercor
  • Location: United States
  • Language: Punjabi

This role consists of evaluating AI models generating music and lyrics in Punjabi. You will analyze compositions generated by AI, compare lyrics with published songs, and rate their quality, creativity and linguistic authenticity. The profile sought is a musician, songwriter or music journalist with expertise in the Punjabi music scene.

Posted August 12, 2026

OpenRecently verified$15/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: Mercor
  • Location: United States
  • Language: Thai

You are an experienced musician who is fluent in Thai: this job consists of evaluating generative AI music models by analyzing generated lyrics, their quality, their authenticity and their compliance with given instructions. Working in Thai and English, you will rate the results according to detailed criteria and compare the lyrics with published songs to identify similarities.

Posted August 12, 2026

OpenRecently verified$18/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States
  • Level: Expert

This role invites a preclinical research expert to advise an AI and foundation models platform applied to drug discovery. You will bring your experience in developing ADC (antibody-drug conjugate) or bispecific molecules to guide AI tool design, by annotating scientific data and evaluating key development decisions. The role merges your laboratory expertise with a growing understanding of how scientific data drives next-generation AI models.

Posted August 12, 2026

OpenRecently verified$60-100/hr

Referral link: we may earn a fee. Apply without it

  • Writing, Creative & Design
  • Platform: micro1
  • Location: 58 countries

You will evaluate written content and responses generated by AI to verify their accuracy, consistency, and overall quality. You will provide detailed written feedback and document your analyses according to the guidelines provided. This role is intended for native English speakers with recognized expertise in professional writing, editing, and proofreading.

Posted August 12, 2026

OpenRecently verified$40-50/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries

This role involves leveraging your tax expertise to contribute to training next-generation AI systems. You will analyze US federal and state sales and use tax laws, precisely translating legal requirements into testable criteria that AI must meet. The ideal profile combines recognized accounting qualification with the ability to break down legal texts into discrete and verifiable elements.

Posted August 11, 2026

OpenRecently verified$20-30/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Legal
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role invites experienced corporate lawyers to contribute to the training and evaluation of AI models specialized in contract analysis. You will participate in contract negotiation and review exercises, evaluate the responses of AI systems, and help refine contract examination tools. This role is intended for seasoned technology law attorneys who wish to shape the next generation of AI-powered legal solutions.

Posted August 11, 2026

OpenRecently verified$100-130/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Legal
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role invites legal experts specialized in technology law to participate in the evaluation and improvement of AI models designed for legal practice. You will contribute to the training and evaluation of contract review tools by providing expert feedback and creating rigorous evaluation frameworks. This role is suited to lawyers with solid experience in negotiating and drafting technology contracts who wish to shape the future of intelligent legal solutions.

Posted August 11, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Engineering
  • Platform: micro1
  • Location: 58 countries
  • Level: Expert

This role consists of designing and evaluating advanced technical problems in engineering fields such as aerodynamics, mechanical design, or electronics, in order to train AI systems. You will need to analyze responses generated by models, apply your engineering judgment, and provide structured feedback to improve their learning and reasoning capabilities.

Posted August 11, 2026

OpenRecently verified$60-110/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States

You will bring your expertise in atomistic modeling and computational chemistry to a cutting-edge AI laboratory in materials science. Your role consists of generating and evaluating high-quality scientific data for training AI models, by applying your in-depth knowledge of surface simulation, adsorption and reaction kinetics. You will contribute directly to improving the scientific reasoning of models on chemical processes and material properties.

Posted August 6, 2026

OpenRecently verified$84/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Terac
  • Location: United States

This ongoing project hires advanced STEM experts to write graduate-level and PhD-level problems with exact, verifiable answers, helping AI systems improve at rigorous scientific reasoning. Participants create original problems in their own specialty, write detailed step-by-step solutions, and sometimes review AI-generated answers to find reasoning errors. It is aimed at PhD candidates and holders, postdoctoral researchers, and faculty in biology, chemistry, physics, mathematics, and selected engineering fields.

Posted August 6, 2026

OpenRecently verifiedPay not disclosed
  • Health & Medicine
  • Platform: Mercor
  • Location: United States
  • Level: Expert

You will contribute to AI expert evaluation by creating and critiquing analysis frameworks for commercial drugs and pharmaceutical development programs. The role focuses primarily on patient population sizing (prevalence, incidence, diagnostic-therapeutic pathway) for different indications, ensuring the quality and reliability of estimates obtained from real-world data.

Posted August 5, 2026

OpenRecently verified$150-175/hr

Referral link: we may earn a fee. Apply without it

  • Software Engineering
  • Platform: Mercor
  • Location: United States

This role involves evaluating advanced AI systems on their ability to analyze complex cybersecurity scenarios, identify vulnerabilities and propose exploits. The ideal candidate is an experienced offensive security researcher with publicly recognized technical contributions and deep expertise in vulnerability research and security engineering.

Posted August 4, 2026

OpenRecently verified$200-250/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

Mercor is seeking Atomic Layer Deposition (ALD) experts and thin film process specialists to support a cutting-edge AI research laboratory developing models for semiconductors and physical sciences. You will contribute your specialized expertise to generate, structure, and evaluate the scientific data used for training and evaluating these advanced models.

Posted August 1, 2026

OpenRecently verified$84/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

Mercor is recruiting scientists and engineers specialized in inorganic synthesis, materials characterization, superconductors and semiconductors to support a cutting-edge AI research laboratory. You will contribute to generating, structuring and evaluating high-quality scientific data intended for model training, by applying your deep expertise to complex problems in materials physics.

Posted August 1, 2026

OpenRecently verified$84/hr

Referral link: we may earn a fee. Apply without it

  • Software Engineering
  • Platform: Mercor
  • Location: United States

This role is for engineers who have built and deployed search systems, particularly in the context of AI agents and large language models. You will be evaluated on your hands-on experience in relevance optimization, impact assessment, and the technical trade-offs you have had to make in real-world systems.

Posted August 1, 2026

OpenRecently verified$80-150/hr

Referral link: we may earn a fee. Apply without it

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

679 AI Evaluation AI training jobs are open on SideHustler today. 26 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 8, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 637 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (139), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.