AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

679
open jobs
26
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 637 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 8, 2026.

481 to 500 of 679 jobs
  • Engineering
  • Platform: Mercor
  • Location: United States

You are an engineer specialized in propulsion and pyrotechnic systems. You will write calibrated technical questions to test an AI model's ability to distinguish legitimate requests from bypass attempts, then evaluate its responses and document the technical reasoning that justifies each judgment.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

Join a team of radiation safety experts to test the robustness of AI models against requests related to dual-use applications. You will write nuanced evaluation questions, assess model responses against strict safety criteria, and document your technical judgment to distinguish legitimate requests from dangerous ones.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

Mercor is looking for nuclear and safeguards experts to test the robustness of AI models against dual-use requests. You will write targeted prompts on the knife's edge between legitimate professional questions and dangerous requests, then evaluate the model's responses against a defined policy and write reference answers. This role is for practitioners with concrete experience in nuclear material verification and detecting diversion attempts.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

You will contribute to evaluating frontier AI models as a source security specialist. Your role is to write prompts that test the model's ability to distinguish legitimate professional questions from potentially dangerous requests, then evaluate its responses against a defined policy standard. This work requires practical expertise in Category 1 and 2 radioactive source security, as well as strong technical writing skills to justify your assessments.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

You will contribute to evaluating AI models as a specialist chemist in energetic materials, by writing technical requests that test the model's ability to distinguish legitimate questions from dangerous requests. You will need to master the fine line between civilian applications and misuse to design realistic scenarios and evaluate the model's responses.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

You will write questions and scenarios to test an AI model's ability to correctly judge the potential for misuse of technical requests in chemistry, distinguishing legitimate questions from dangerous requests. This role requires practical expertise as a synthetic chemist or process engineer with experience scaling up reactions, and the ability to write clear technical justifications for non-specialists.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Language: French, Malay, Russian, Chinese, Japanese

This role involves supporting the development of an AI system by bringing your mental health expertise to it. You will analyze and evaluate psychological content generated by the AI, develop clinical datasets, and provide detailed feedback to improve the model's responses in therapeutic contexts. You must have fluent command of English and at least one other language from a list of more than fifteen languages.

Posted September 14, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries

This role calls upon investment banking experts to contribute to training AI systems in the field of financial analysis. You will put your experience in mergers and acquisitions to use creating high-quality training data, by performing comparable analyses, building financial models, and evaluating the quality of analytical work.

Posted September 14, 2026

OpenRecently verified$175-300/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: micro1
  • Location: 58 countries

This role enables you to analyze and optimize diverse codebases, as well as design programming tasks to evaluate generative AI models. It is intended for experts in algorithmics and competitive programming problem-solving, capable of working autonomously and providing high-quality technical feedback.

Posted September 14, 2026

OpenRecently verified$50-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Engineering
  • Platform: micro1
  • Location: 58 countries

This robotics training role consists of operating a robotic arm manipulator (UMI gripper) in a laboratory to perform repetitive manipulation tasks for the training of robotic AI models. You must execute detailed instructions with precision, document results, and report any anomalies encountered. No prior AI experience is required, but excellent manual dexterity, reliability, and strong concentration skills are essential.

Posted September 14, 2026

OpenRecently verified$30/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Writing, Creative & Design
  • Platform: micro1
  • Location: 58 countries

You will annotate and evaluate videos and audio recordings to train AI systems, verifying the accuracy of generated descriptions and providing detailed feedback. This role is suited to specialists with solid domain expertise in analyzing visual and audio content, without requiring prior experience in artificial intelligence.

Posted September 14, 2026

OpenRecently verified$8-10/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Engineering
  • Platform: Mercor
  • Location: United States

You put your expertise in mechanical design and technical drawing at the service of an annotation and evaluation job. You work directly in AutoCAD Mechanical on Windows to capture and document industrial drawing processes according to engineering standards. This role is for mechanical designers, drafters and manufacturing engineers with hands-on experience with this software.

Posted September 13, 2026

OpenRecently verified$45-55/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries

This role is for mental health professionals specialized in trauma of children and adolescents who are abuse survivors. You will contribute to training AI systems by bringing your clinical expertise to ensure models provide sensitive and accurate guidance to young survivors. No prior AI experience is required - only your domain knowledge matters.

Posted September 13, 2026

OpenRecently verified$100-200/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: Mercor
  • Location: United States

This role is for Excel experts with solid experience in financial modeling. You will contribute to improving the training of next-generation AI models by designing realistic Excel tasks, writing precise solutions, and evaluating the quality of financial data produced. The role suits professionals from private equity, investment banking, hedge funds, equity research, or accounting.

Posted September 12, 2026

OpenRecently verified$70-100/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States

This role involves leveraging your advanced Excel expertise to improve the training of generative AI models. You will design realistic Excel tasks drawn from your professional experience, provide structured solutions, and evaluate the work of other experts to ensure the quality of training data. The ideal profile is a versatile generalist who has applied Excel at a high level across multiple sectors and business functions.

Posted September 12, 2026

OpenRecently verified$70-100/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States

You will join a team as a quality analyst to evaluate the accuracy and consistency of work produced. You will examine results against defined standards, identify errors and inconsistencies, and provide structured feedback to continuously improve quality processes.

Posted September 12, 2026

OpenRecently verified$20-50/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United Kingdom

This role is for molecular biology experts capable of independently designing nucleic acid sequences. You will contribute to improving AI models by evaluating their reasoning in the field, designing complex tasks, and writing reference solutions that will be used to train models. You will also establish quality criteria and guidelines for judging sequence design.

Posted September 11, 2026

OpenRecently verified$70-105/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

You will join a cutting-edge generative artificial intelligence team to contribute to the development of foundational AI models. As a molecular biology expert, you will design complex tasks, validate the quality of training data, and establish evaluation criteria to optimize the biological reasoning of AI models. This role is intended for researchers with solid hands-on experience in nucleic acid construct design and high-level scientific output.

Posted September 11, 2026

OpenRecently verified$70-105/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: micro1
  • Location: 58 countries
  • Language: Chinese

This role involves evaluating the quality and relevance of texts and statements in Cantonese to train generalist AI systems. As a native Cantonese speaker, you will use proprietary tools to analyze linguistic data and provide detailed feedback aimed at improving natural language processing models.

Posted September 11, 2026

OpenRecently verified$30-40/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Business, Consulting & Operations
  • Platform: micro1
  • Location: 58 countries

This role consists of evaluating and scoring responses generated by AI systems on various types of professional content (specifications, release notes, communications). You will analyze writing quality, compliance with instructions, and suitability for the target audience, providing detailed justifications. This role is aimed at experienced product managers or product owners who know how to evaluate technical and business content.

Posted September 11, 2026

OpenRecently verified$90-140/hr

Referral link to micro1's job list: search for this role there. Open this exact job

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

679 AI Evaluation AI training jobs are open on SideHustler today. 26 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 8, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 637 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (139), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.