1234 open jobs

Remote AI training jobs

Open roles in AI training, model evaluation, RLHF and data labeling for domain experts. Search, filter, then apply directly on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac.

13 jobs
  • Software Engineering
  • Platform: Terac
  • Location: United States

This remote study asks practicing software engineers to review programming tasks and the evaluation harnesses built to test AI agents on them. You will check whether the tasks, test cases and environment structure reflect realistic software engineering work. It suits engineers with solid experience building, testing and reviewing complex systems.

Posted October 2, 2026

  • Software Engineering
  • Platform: Terac
  • Location: Argentina

This paid study asks experienced software engineers to review proposed programming tasks and the harnesses that test AI agents, checking their logic, structure and realism. The work takes place in a screen-shared session that mixes code review, technical discussion and direct feedback on task design. It is aimed at engineers who have already built, reviewed or tested evaluation harnesses or coding tasks.

Posted October 1, 2026

  • Software Engineering
  • Platform: Terac
  • Location: India

This is a paid research interview in which experienced software engineers review proposed programming tasks and their evaluation environments. Participants assess how realistic, difficult and well structured the challenges are, and they suggest ways to make the test suites more robust. It suits engineers who already build or review complex codebases and know how test harnesses work.

Posted October 1, 2026

  • Business, Consulting & Operations
  • Platform: Terac
  • Location: United States

This paid study asks operations professionals who manage internal data at architecture firms to take part in a remote video interview. Participants explain how they turn messy operational data into structured formats used for post-training evaluation, and they review hypothetical data transformation scenarios. It is aimed at data operations leads, BIM managers, and workflow specialists with direct hands-on experience.

Posted September 27, 2026

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This remote freelance role involves labeling and reviewing data to help improve artificial intelligence models. It suits people with prior experience in data annotation, data labeling, or AI model evaluation who can work independently on a flexible schedule. Tasks range from categorizing text and tagging entities to judging the quality and safety of AI-generated responses.

Posted September 25, 2026

  • Business, Consulting & Operations
  • Platform: Terac
  • Location: United States

This paid study invites experienced management consultants and strategy professionals to evaluate AI-generated business strategy and advisory content. Participants work remotely, marking responses for factual accuracy, logical structure and professional tone. Their written feedback is used to help improve the reliability of these business analysis tools.

Posted September 18, 2026

OpenRecently verifiedPay not disclosed
  • Finance & Accounting
  • Platform: Terac
  • Location: United States

This paid study asks finance professionals to evaluate AI-generated documents and data points tied to transaction finance. Reviewers rank model outputs by accuracy and relevance, and annotate errors or points where the model's reasoning fails. It is aimed at people with hands-on experience in M&A, investment banking, corporate development, or financial analysis, and the tasks are completed remotely at one's own pace.

Posted September 18, 2026

OpenRecently verifiedPay not disclosed
  • Health & Medicine
  • Platform: Terac
  • Location: United States

This paid study asks practicing dermatologists to judge how well AI diagnostic tools perform on skin images and clinical records. Participants review de-identified cases and check the automated findings against established reference standards. It suits clinicians who routinely diagnose skin conditions and can assess how reliable generated diagnoses are.

Posted September 10, 2026

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This study invites machine learning engineers and AI researchers to build test scenarios on a remote reinforcement learning platform. Participants set environmental parameters, define spatial constraints and agent interaction rules, run preliminary agent tests, and then share usability feedback in an interview. It is aimed at people with hands-on experience in simulation design and RL environments.

Posted September 3, 2026

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This project asks evaluators to examine how a digital shopping assistant answers real e-commerce questions and to locate where its reasoning or product suggestions fall short. The work is remote and ongoing, and it is aimed at people with quality assurance, data evaluation, or e-commerce backgrounds who can build structured evaluation frameworks.

Posted September 2, 2026

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This is a paid one-hour trial for people who know a professional or personal workflow well. Participants turn that workflow into a demanding prompt, test it in ChatGPT to find where the model fails, and make the prompt harder when the model succeeds. The work suits domain experts and power users who can judge AI outputs quickly and who are open to ongoing evaluation work.

Posted August 25, 2026

OpenRecently verifiedPay not disclosed
  • Finance & Accounting
  • Platform: Terac
  • Location: United States

This remote project asks experienced accounting and financial planning professionals to write complex financial scenarios that will help AI systems handle sophisticated accounting tasks. Participants design edge cases, build evaluation rubrics for planning and analysis work, and peer-review scenarios written by other experts. It is aimed at CPAs, controllers, FP&A directors and senior financial analysts with deep technical expertise.

Posted August 19, 2026

  • Science & Research
  • Platform: Terac
  • Location: United States

This ongoing project hires advanced STEM experts to write graduate-level and PhD-level problems with exact, verifiable answers, helping AI systems improve at rigorous scientific reasoning. Participants create original problems in their own specialty, write detailed step-by-step solutions, and sometimes review AI-generated answers to find reasoning errors. It is aimed at PhD candidates and holders, postdoctoral researchers, and faculty in biology, chemistry, physics, mathematics, and selected engineering fields.

Posted August 6, 2026

OpenRecently verifiedPay not disclosed

FAQ

Questions about AI training jobs

What kinds of jobs are listed here?

AI training, model evaluation, RLHF, red teaming, data labeling and expert review roles, in fields from software and data science to medicine, law, finance, languages and science.

Are these jobs remote?

They are done online. Some are limited to residents of certain countries: when the platform publishes that restriction, the listing shows it.

Do I need prior experience in AI?

Usually not. Most roles ask for professional experience in your own field. Each listing states the experience the platform requires.

How is the work paid?

By the hour or by the task, by the platform that hires you. We show the pay only when the platform publishes it. It is not a guarantee of income or hours.

How do I apply?

Apply now opens the official posting on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac, where you complete the application. You need no account here. See how it works.