AI Evaluation AI training jobs

Open AI training and model evaluation roles that call for AI Evaluation expertise, gathered from 10 platforms. Part of Data & Machine Learning.

679
open jobs
26
posted this week
$75/hr
median published rate
$400/hr
highest published rate
10
platforms hiring

Pay figures cover the 637 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 8, 2026.

41 to 60 of 679 jobs
  • Data & Machine Learning
  • Platform: Alignerr

This role involves creating and evaluating biology problems to test and improve AI models' scientific reasoning capabilities. Biology masters-level experts will design rigorous test cases, assess AI outputs for accuracy, and provide feedback to enhance model performance across various biology domains.

  • Science & Research
  • Platform: Alignerr

This role offers biologists with advanced degrees the opportunity to contribute to AI model training by developing rigorous biology problems, evaluating AI-generated scientific content, and ensuring scientific accuracy in AI reasoning. Remote and flexible, it leverages deep domain expertise to shape how AI systems approach complex problems in genetics, biotechnology, and bioinformatics.

  • Science & Research
  • Platform: Alignerr

This role involves collaborating with an AI training platform to develop and evaluate biology problems, stress-test AI reasoning on genetics and biotechnology topics, and ensure scientific accuracy in AI-generated biological explanations. It is designed for advanced researchers and scientists with deep domain expertise who want flexible, intellectually challenging work without requiring prior AI experience.

  • Data & Machine Learning
  • Platform: Alignerr

This role involves leading data governance initiatives for biotech and clinical trial data to ensure compliance, accuracy, and readiness for AI model training. The position requires bridging scientific, technical, and regulatory teams to establish policies that protect sensitive data while enabling innovation in life sciences and regulatory submissions.

  • Business, Consulting & Operations
  • Platform: Alignerr

This role involves helping train and evaluate AI models for business applications by designing complex case studies, authoring gold-standard solutions, and auditing AI reasoning across corporate finance and strategy scenarios. The position is designed for MBA-qualified professionals who can stress-test AI systems on real-world business challenges and ensure outputs meet executive-level standards.

  • Business, Consulting & Operations
  • Platform: Alignerr
  • Level: Intermediate

This role involves leading and mentoring teams of business analysts who create high-quality training content for AI models, with a focus on business, finance, strategy, and operations domains. The position requires strong team management and business analysis expertise to oversee content quality, provide performance feedback, and serve as the stakeholder point of contact for remote contract work.

  • Business, Consulting & Operations
  • Platform: Alignerr
  • Level: Intermediate

This role involves designing and evaluating business content used in AI training datasets. Experienced postsecondary business instructors will review AI-generated materials, create realistic scenarios, and provide feedback to improve how AI systems understand and communicate business concepts. The position is fully remote and flexible, suited for educators who want to influence AI development.

  • Software Engineering
  • Platform: Alignerr
  • Level: Expert

A senior C# engineer will design and build high-performance data pipelines, annotation tooling, and evaluation systems for AI model training. The role involves full-stack backend development, system optimization, and collaboration with data and research teams to support production infrastructure for AI labs.

  • Engineering
  • Platform: Alignerr

This role involves designing small, self-contained computational fluid dynamics tasks that train and evaluate AI models on engineering reasoning. CFD engineers create problem statements, scoring mechanisms, and reference solutions to teach AI systems about aerodynamics, boundary conditions, and flow physics, without requiring machine learning expertise.

  • Science & Research
  • Platform: Alignerr

This role involves annotating and evaluating chemistry content for AI model training. A PhD chemist will label complex chemistry questions, map conceptual relationships, review AI-generated answers, and develop datasets across multiple chemistry subfields to improve frontier AI models.

  • Science & Research
  • Platform: Alignerr

This role involves designing, solving, and evaluating advanced chemistry problems to train AI models. Masters and PhD-level chemists will leverage their domain expertise to create rigorous scientific content that enhances AI reasoning capabilities across organic, inorganic, physical, or computational chemistry.

  • Science & Research
  • Platform: Alignerr

Chemistry experts with PhD-level qualifications design rigorous chemistry problems and evaluate AI model responses for scientific accuracy and reasoning quality. This remote contract role involves assessing how AI systems understand complex chemistry concepts across multiple domains and providing detailed feedback to improve model performance.

  • Data & Machine Learning
  • Platform: Alignerr

Chemistry experts with a Master's degree help train AI systems to reason through complex scientific problems by designing chemistry problems, evaluating AI outputs, and providing feedback. This remote position suits chemistry specialists who want to apply their expertise to improve how AI models handle scientific content.

  • Science & Research
  • Platform: Alignerr

This role involves designing, solving, and evaluating complex chemistry problems to train AI models in scientific reasoning. Chemistry researchers with advanced degrees will apply their domain expertise to create rigorous problem sets and evaluate how AI systems understand chemistry concepts, working remotely on a flexible schedule.

  • Science & Research
  • Platform: Alignerr

This role involves designing, solving, and evaluating complex chemistry problems to train AI models. Chemistry experts with advanced degrees will apply their scientific knowledge to create rigorous problem sets and provide feedback on AI-generated scientific outputs, working remotely on a flexible schedule.

  • Data & Machine Learning
  • Platform: Alignerr

This role involves evaluating and improving AI systems that handle advanced chemistry problems by designing rigorous test cases and assessing the accuracy and reasoning quality of AI-generated solutions. Chemistry experts with Master's-level knowledge will work remotely to help AI systems better understand and reason through chemistry across multiple subdisciplines.

  • Writing, Creative & Design
  • Platform: Alignerr

Cinematography professionals and advanced film students analyze film clips to train AI systems to understand visual storytelling. You will break down camera work, shot composition, framing, and scene structure into detailed, structured annotations that help AI perceive and describe cinema.

  • Writing, Creative & Design
  • Platform: Alignerr

This role involves analyzing short video clips and writing precise cinematographic descriptions to help train AI video generation models. The position is ideal for film professionals who can articulate visual techniques using accurate cinema terminology and structured prompts.

  • Data & Machine Learning
  • Platform: Alignerr

This role involves developing environmental science problems and evaluating AI model outputs for accuracy in climate, ecology, and sustainability domains. Climate Data Specialists with advanced environmental science degrees will design complex scientific scenarios and assess AI-generated content for scientific rigor and clarity in support of AI model training.

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about AI Evaluation AI training jobs

How many AI Evaluation AI training jobs are open right now?

679 AI Evaluation AI training jobs are open on SideHustler today. 26 were posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 8, 2026.

How much do AI Evaluation AI training jobs pay?

Among the 637 open roles that publish an hourly rate in USD, pay runs from $8 to $400 per hour, with a median of $75. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer AI Evaluation AI training jobs?

Alignerr (253), Mercor (196), micro1 (139), DataAnnotation (47), Terac (13), Vetto (11), Handshake AI (7), Mindrift (6), xAI (4) and Ethos (3). You apply on the platform itself, which handles screening, contracts and payment.

What skills do AI Evaluation AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Data Annotation, Content Annotation, Code Quality & Review, Multilingual Expertise. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.