1234 open jobs

Remote AI training jobs

Open roles in AI training, model evaluation, RLHF and data labeling for domain experts. Search, filter, then apply directly on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac.

"evaluation"×Clear all
221 to 240 of 319 jobs
  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States

You will join a team as a quality analyst to evaluate the accuracy and consistency of work produced. You will examine results against defined standards, identify errors and inconsistencies, and provide structured feedback to continuously improve quality processes.

Posted September 12, 2026

OpenRecently verified$20-50/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

You will join a cutting-edge generative artificial intelligence team to contribute to the development of foundational AI models. As a molecular biology expert, you will design complex tasks, validate the quality of training data, and establish evaluation criteria to optimize the biological reasoning of AI models. This role is intended for researchers with solid hands-on experience in nucleic acid construct design and high-level scientific output.

Posted September 11, 2026

OpenRecently verified$70-105/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: micro1
  • Location: 58 countries

This role consists of evaluating and scoring responses generated by AI systems on various types of professional content (specifications, release notes, communications). You will analyze writing quality, compliance with instructions, and suitability for the target audience, providing detailed justifications. This role is aimed at experienced product managers or product owners who know how to evaluate technical and business content.

Posted September 11, 2026

OpenRecently verified$90-140/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: Terac
  • Location: United States

This paid study asks practicing dermatologists to judge how well AI diagnostic tools perform on skin images and clinical records. Participants review de-identified cases and check the automated findings against established reference standards. It suits clinicians who routinely diagnose skin conditions and can assess how reliable generated diagnoses are.

Posted September 10, 2026

  • Science & Research
  • Platform: Mercor
  • Location: United States

You are a chemistry expert with practical experience in synthesis, analytical chemistry, or chemical safety. You participate in evaluating AI models by writing calibrated test prompts across three risk levels, then assessing model responses against an established safety framework. This role requires both advanced technical expertise and the ability to document your judgments in an accessible way.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Writing, Creative & Design
  • Platform: micro1
  • Location: 58 countries

You will contribute to training AI models by evaluating and selecting images according to rigorous aesthetic criteria. Your visual expertise will enable you to judge the quality of composition, lighting, color and originality, while documenting your analyses to calibrate quality standards with other specialists.

Posted September 8, 2026

OpenRecently verified$21-70/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States
  • Language: Croatian

This role involves evaluating and improving the safety of AI models by analyzing their behavior on sensitive topics in Croatian. You will write expert prompts, classify content according to structured guidelines, and identify adversarial formulations, as an English-Croatian bilingual speaker with no prior AI experience required.

Posted September 4, 2026

OpenRecently verified$38-42/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: Mercor
  • Location: Belgium
  • Language: Dutch

This role asks bilingual English and Dutch speakers to test how AI models respond to sensitive topics in Dutch as used in Belgium. The work combines language fluency with cultural judgment, and no prior AI experience is needed because the workflow is taught on the job.

Posted September 4, 2026

OpenRecently verified$48-52/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Turkish

You will participate in improving AI model safety by evaluating how they handle sensitive subjects in Turkish. Your linguistic and cultural judgments will help identify and strengthen weaknesses in these systems when facing delicate content. No prior AI experience is required.

Posted September 4, 2026

OpenRecently verified$23-27/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

This role involves evaluating and annotating clinical documents as well as notes generated by medical AI systems, applying your expertise in coding and clinical documentation. You will participate in a multilingual team responsible for ensuring the accuracy and reliability of AI tools in healthcare settings. The ideal candidate must possess recognized medical certification and be proficient in English as well as at least one other language at an advanced level.

Posted September 3, 2026

OpenRecently verified$45-65/hr

Referral link: we may earn a fee. Apply without it

  • Engineering
  • Platform: micro1
  • Location: 58 countries

This role involves training AI models by leveraging your expertise in CAD. You will evaluate content generated by AI in manufacturing workflows, provide detailed feedback on manufacturing processes, and develop evaluation frameworks to measure accuracy and efficiency. The role requires strong programming skills, particularly with Python and coding agents.

Posted September 3, 2026

OpenRecently verified$80-150/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Health & Medicine
  • Platform: Mercor
  • Location: United States
  • Language: Spanish, Chinese, Filipino, Russian, Haitian Creole
  • Level: Intermediate

This role is for practicing inpatient hospitalist physicians to evaluate and annotate clinical documentation generated by medical AI systems. You will review patient records, identify gaps and inaccuracies in notes produced by AI, and provide structured feedback to engineering teams. The role requires advanced proficiency in at least one language other than English.

Posted September 2, 2026

OpenRecently verified$170/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States
  • Language: Spanish, Chinese, Filipino, Russian, Haitian Creole
  • Level: Intermediate

You are a general practitioner or internist currently practicing in an outpatient setting and proficient in English plus another language at an advanced level. You annotate and evaluate clinical notes generated by AI, flagging omissions and inaccuracies to improve clinical decision support tools. This remote and part-time role allows you to leverage your expertise in medical documentation with a specialized AI health product team.

Posted September 2, 2026

OpenRecently verified$170-190/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This project asks evaluators to examine how a digital shopping assistant answers real e-commerce questions and to locate where its reasoning or product suggestions fall short. The work is remote and ongoing, and it is aimed at people with quality assurance, data evaluation, or e-commerce backgrounds who can build structured evaluation frameworks.

Posted September 2, 2026

  • Health & Medicine
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role asks you to evaluate and analyze complex cases of eating disorders in adolescents by applying your clinical knowledge to contribute to the training of AI systems. You will produce detailed evaluations, diagnoses based on validated criteria, and treatment recommendations adapted for non-specialists, in a fully remote work setting.

Posted September 1, 2026

OpenRecently verified$45-95/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Science & Research
  • Platform: micro1
  • Location: 58 countries

You will contribute to an AI training project by solving, validating, and commenting on complex problems in your scientific field of expertise. The role combines rigorous scientific analysis, Python programming for calculations and simulations, and critical evaluation of solutions generated by AI.

Posted September 1, 2026

OpenRecently verified$80-150/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Science & Research
  • Platform: Mercor
  • Location: United States

This role involves designing realistic evaluation challenges for AI models applied to drug discovery. As an expert in chemistry and molecular biology, you will build complex research scenarios based on real and ambiguous data, where AI must demonstrate genuine reasoning capability rather than simple pattern association. You will be responsible for the entire process: challenge design, data corpus assembly, response evaluation, and iterative revision.

Posted August 31, 2026

OpenRecently verified$60-90/hr

Referral link: we may earn a fee. Apply without it

  • Engineering
  • Platform: Mercor
  • Location: United States

This evaluation role involves designing simulations and technical specifications to test the ability of advanced AI models to reason from first principles in your area of expertise in engineering. You will define complex design problems (controllers, circuits, geometries) that the AI must solve by demonstrating reasoning capability rather than simple data retrieval.

Posted August 31, 2026

OpenRecently verified$60-90/hr

Referral link: we may earn a fee. Apply without it

  • Software Engineering
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

You will evaluate the quality and correctness of AI-assisted software development traces used to train and assess models at a leading AI laboratory. You must judge the accuracy of coding sessions, the consistency of workflows, and the quality of reasoning, then provide structured written feedback according to an evaluation grid.

Posted August 28, 2026

OpenRecently verified$70-90/hr

Referral link: we may earn a fee. Apply without it

  • Software Engineering
  • Platform: Mercor
  • Location: United States
  • Level: Intermediate

You will assess the quality and architectural robustness of serverless AWS tasks and infrastructure-as-code designed for AI model training. Your role consists of analyzing multi-service designs, verifying the fidelity of infrastructure code, and providing detailed feedback based on evaluation grids. This profile is aimed at cloud engineers with solid practical experience in AWS serverless services.

Posted August 28, 2026

OpenRecently verified$70-90/hr

Referral link: we may earn a fee. Apply without it

FAQ

Questions about AI training jobs

What kinds of jobs are listed here?

AI training, model evaluation, RLHF, red teaming, data labeling and expert review roles, in fields from software and data science to medicine, law, finance, languages and science.

Are these jobs remote?

They are done online. Some are limited to residents of certain countries: when the platform publishes that restriction, the listing shows it.

Do I need prior experience in AI?

Usually not. Most roles ask for professional experience in your own field. Each listing states the experience the platform requires.

How is the work paid?

By the hour or by the task, by the platform that hires you. We show the pay only when the platform publishes it. It is not a guarantee of income or hours.

How do I apply?

Apply now opens the official posting on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac, where you complete the application. You need no account here. See how it works.