AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

679
open jobs
26
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 637 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 8, 2026.

541 to 560 of 679 jobs
  • Business, Consulting & Operations
  • Platform: micro1
  • Location: 58 countries
  • Level: Entry level

This role consists of evaluating and improving AI model responses to strategy and business consulting problems. You will leverage your consulting firm experience to design quality prompts, write reference responses and provide detailed feedback, applying structured analytical frameworks to identify reasoning gaps.

Posted August 27, 2026

OpenRecently verified$245-280/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role requires a private equity specialist to design authentic tasks and evaluate results generated by AI systems trained on industry workflows. You will critique financial models, memorandums and deal analyses, ensuring methodological rigor and strategic relevance.

Posted August 27, 2026

OpenRecently verified$45-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: micro1
  • Location: 58 countries
  • Level: Entry level

You will evaluate code generated by AI models by examining its correctness, quality, and design. You will also participate in creating complex engineering problems, reference solutions, and test cases to train AI models on realistic software tasks. This role is for senior software engineers who have worked at major technology companies.

Posted August 27, 2026

OpenRecently verified$245-280/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role involves contributing to the creation of high-quality datasets capturing real venture capital practices, in order to train high-performing AI systems. You will apply your industry expertise to design authentic tasks, evaluate contents generated by AI, and refine analysis methodologies. The ideal profile has solid experience in venture capital or as an entrepreneur who has managed funding rounds.

Posted August 27, 2026

OpenRecently verified$45-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Writing, Creative & Design
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

You will annotate and analyze editorial designs, posters and marketing content to train AI systems. As a visual composition specialist, you will document in detail the structure, hierarchy and typographic choices of creative work, using precise design vocabulary and a structured methodology. This role is for experienced graphic designers who can translate their visual expertise into high-quality annotations.

Posted August 26, 2026

OpenRecently verified$20-50/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Software Engineering
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Evaluate the quality and compliance of vulnerability reproduction and remediation exercises designed to train AI models. Your role consists of verifying the accuracy of CVE reproductions, the robustness of fixes, the rigor of verification tests, and the adequacy of Docker environments, while providing detailed and structured feedback.

Posted August 25, 2026

OpenRecently verified$70-90/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States

This role involves conducting structured 20-minute interviews with engineers specialized in GPU kernel development, security research, and ML compilers, to assess their technical expertise and ability to participate in AI evaluation projects. You will provide reasoned recommendations (move forward, hold, or reject) and write concise summaries for the project team. This role is for someone with confirmed experience in recruiting or technical screening of senior profiles.

Posted August 25, 2026

OpenRecently verified$50-60/hr

Referral link: we may earn a fee. Apply without it

  • Finance & Accounting
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

This job consists of training AI models by providing high-quality data on property and casualty insurance and liability underwriting reasoning. You will design realistic underwriting scenarios, evaluate responses generated by AI according to business expertise criteria, and contribute to improving the models' ability to reason about insurability, pricing, and risk selection. This role is intended for experienced underwriters with direct responsibility in risk evaluation and selection.

Posted August 25, 2026

OpenRecently verified$80/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

This role involves creating realistic insurance policy processing scenarios and evaluating an AI model's ability to reason correctly across the entire policy lifecycle. You will design processing instructions, write reference responses, and score AI-generated responses according to precise operational criteria. We are looking for experienced professionals in policy administration, insurance operations management, or portfolio management.

Posted August 25, 2026

OpenRecently verified$80/hr

Referral link: we may earn a fee. Apply without it

  • Finance & Accounting
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

This role involves designing realistic reinsurance and program insurance scenarios, evaluating responses generated by artificial intelligence against established market standards, and contributing to improving the reasoning of cutting-edge models on treaty structures, facultative placements, delegations of authority, and alternative risk transfer mechanisms. It is aimed at experienced professionals in the reinsurance sector, program insurance, brokers, or program administrators.

Posted August 25, 2026

OpenRecently verified$80/hr

Referral link: we may earn a fee. Apply without it

  • Legal
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

This role consists of preparing realistic claims compensation scenarios and evaluating responses generated by AI models according to established insurance industry standards. The ideal candidate is a multiline claims expert with at least two years of professional experience, capable of analyzing insurance policies, identifying gaps in claims handling, and providing detailed feedback to improve AI reasoning.

Posted August 25, 2026

OpenRecently verified$80/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This is a paid one-hour trial for people who know a professional or personal workflow well. Participants turn that workflow into a demanding prompt, test it in ChatGPT to find where the model fails, and make the prompt harder when the model succeeds. The work suits domain experts and power users who can judge AI outputs quickly and who are open to ongoing evaluation work.

Posted August 25, 2026

OpenRecently verifiedPay not disclosed
  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Mercor is seeking experienced computer and information systems managers to evaluate and improve advanced AI systems. You will analyze content and workflows generated by AI in the field of IT management, leveraging your genuine expertise in information systems leadership, technical team coordination, and oversight of enterprise technology operations. This role is open to candidates with a minimum of four years of experience in IT management.

Posted August 24, 2026

OpenRecently verified$70-110/hr

Referral link: we may earn a fee. Apply without it

  • Software Engineering
  • Platform: Mercor
  • Location: United States
  • Level: Expert

This role invites experienced software engineers to evaluate and design migration strategies for complex codebases, particularly in mission-critical environments. You will analyze existing systems to identify safe modernization approaches and verify the quality of solutions generated by AI, while preserving functionality and security.

Posted August 22, 2026

OpenRecently verified$200/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Mercor is recruiting experienced nurse practitioners to evaluate and improve cutting-edge AI systems in healthcare. You will analyze clinical content generated by AI, treatment plans and patient education materials by leveraging your real-world expertise in diagnosing, treating and managing patients with acute and chronic conditions.

Posted August 21, 2026

OpenRecently verified$120-150/hr

Referral link: we may earn a fee. Apply without it

  • Legal
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

You will evaluate compliance content and workflows generated by AI systems, drawing on your experience in regulatory law and law enforcement inspection. This role is for compliance professionals with solid knowledge of legal standards, permits, and oversight procedures.

Posted August 21, 2026

OpenRecently verified$70-110/hr

Referral link: we may earn a fee. Apply without it

  • Finance & Accounting
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Mercor is looking for experienced financial and investment analysts to evaluate the quality and accuracy of financial analysis content generated by AI. You will use your expertise in company valuation, quantitative analysis, and investment recommendations to test and improve these AI systems.

Posted August 21, 2026

OpenRecently verified$70-110/hr

Referral link: we may earn a fee. Apply without it

  • Finance & Accounting
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Mercor is seeking experienced financial managers to evaluate and improve cutting-edge AI systems. You will analyze AI-generated content and workflows in the fields of accounting, banking, credit, and financial planning, leveraging your real-world expertise in managing financial operations. This role is based on assessing the quality, regulatory compliance, and analytical rigor of AI responses.

Posted August 21, 2026

OpenRecently verified$70-110/hr

Referral link: we may earn a fee. Apply without it

  • Legal
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Mercor is recruiting experienced lawyers to evaluate and improve cutting-edge AI systems. You will examine legal content generated by AI, including memorandums, contracts, and case analyses, drawing on your expertise in litigation, legal writing, and client counseling. This role is intended for legal professionals with solid experience in legal practice.

Posted August 21, 2026

OpenRecently verified$70-110/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

Mercor is seeking experienced Medical and Health Services Managers to evaluate and improve AI systems designed for clinical operations. You will analyze content and workflows generated by AI in a healthcare management context, drawing on your hands-on expertise in leading clinical services, personnel management, budgets, and regulatory compliance.

Posted August 21, 2026

OpenRecently verified$100-120/hr

Referral link: we may earn a fee. Apply without it

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

679 AI Evaluation AI training jobs are open on SideHustler today. 26 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 8, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 637 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (139), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.