1234 open jobs

Remote AI training jobs

Open roles in AI training, model evaluation, RLHF and data labeling for domain experts. Search, filter, then apply directly on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac.

"evaluation"×Clear all
81 to 100 of 319 jobs
  • Software Engineering
  • Platform: Alignerr
  • Level: Expert

This contract role asks a senior software engineer to build and scale the backend infrastructure behind AI products, working closely with machine learning engineers and researchers. It suits experienced engineers who can work independently on production-grade services, data pipelines, and APIs.

  • Software Engineering
  • Platform: Alignerr
  • Level: Expert

This contract role is for senior engineers who build the backend systems behind AI model development. The work covers high-performance pipelines, annotation tooling, and evaluation infrastructure that support model training and benchmarking. It is aimed at engineers with several years of production experience in Go, Rust, Python, or C++ who are fully remote and available 20 to 40 hours per week.

  • Software Engineering
  • Platform: Alignerr
  • Level: Intermediate

This is a remote contract role building internal C tooling that supports data pipelines, annotation systems, and evaluation workflows for AI training. It is aimed at senior engineers with several years of production C experience who can also work across full-stack and interoperability tasks.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Swedish

This remote, flexible contract role asks Swedish speakers to review text produced by AI, checking whether it is accurate, reads naturally and suits the cultural expectations of real Swedish readers. No prior AI experience is needed, since the work relies on a strong command of Swedish and careful judgment of written quality.

  • Software Engineering
  • Platform: Alignerr
  • Level: Intermediate

This is a remote hourly contract for a senior Rust engineer who will design and maintain high-performance data pipelines, annotation tooling, and evaluation systems used to train AI models. The work is production-focused, involves backend and tooling services, and suits engineers who can work independently with research and data teams.

  • Software Engineering
  • Platform: Alignerr
  • Level: Expert

This is a remote hourly contract for senior systems engineers who will build and tune the C++ infrastructure behind AI data pipelines, annotation tools and model evaluation. The work centers on performance, reliability and production fixes, carried out alongside research and engineering teams. It suits experienced C++ developers with a systems background who want to work on the infrastructure used to train and ship AI models.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Filipino

This remote contract role asks Tagalog speakers to review AI-generated text and original content for accuracy, naturalness and cultural fit. It is aimed at people with a strong command of contemporary Tagalog who can give detailed, structured feedback on language quality, and no prior AI experience is required.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Turkish

This remote, hourly contract asks Turkish speakers to check AI-generated text for accuracy, naturalness, and cultural fit. Work is done independently on a flexible schedule, with no prior AI experience needed. It suits people with native or near-native Turkish who can give structured, detailed feedback on written language.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Urdu

This contract role asks Urdu speakers to assess and refine text generated by AI systems, checking that it is accurate, reads naturally, and suits real Urdu readers. It is aimed at people with native or near-native Urdu and a careful, methodical eye for written quality, and no prior AI background is needed. The work is remote and performed asynchronously on the contractor's own schedule.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: French

This remote contract role asks French voice professionals to record speech samples that help AI systems sound more natural and expressive. It also involves judging AI-generated French audio and giving feedback on its quality. It is aimed at working voice actors and voice-over artists who have a professional studio setup and no prior AI background.

OpenRecently verified$170-200/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Arabic

This role involves evaluating and improving Arabic translations and AI-generated text by assessing fluency, tone, and cultural authenticity. Native or near-native Arabic speakers will teach language models to produce more natural, human-like Arabic output by rating translations, identifying gaps that automated systems miss, and rewriting model outputs to match native speaker standards.

Talent poolRecently verified$25-40/hr
  • Legal
  • Platform: DataAnnotation

This role evaluates how AI models handle audit procedures, internal controls, and assurance judgments by testing their reasoning, identifying gaps in professional skepticism, and writing correct audit approaches. It is for experienced auditors who can assess whether AI-generated evidence actually supports audit opinions.

Talent poolRecently verified$40-125/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Bangla

A Bengali specialist evaluates and corrects AI-generated Bengali text, focusing on accurate pronoun usage, regional variations, and natural phrasing. The role involves reviewing translations, scoring model outputs, and providing native-speaker feedback to improve AI language understanding.

Talent poolRecently verified$25-40/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Catalan

Evaluate AI model performance on Catalan translation and localization tasks, assessing translation quality and cultural appropriateness. The role involves reviewing AI-generated outputs, identifying errors, and providing corrections and explanations to improve model training.

Talent poolRecently verified$25-40/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Chinese

A Chinese language specialist who evaluates and corrects AI-generated Mandarin text for naturalness, cultural appropriateness, and native fluency. The role involves rating model outputs, identifying errors in measure words and idioms, and rewriting content to sound authentically Chinese rather than translated.

Talent poolRecently verified$25-40/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Czech

This role involves evaluating AI model outputs on Czech translation and localization tasks, identifying errors, and providing corrected versions to improve AI training. The position is suited for Czech language specialists who can assess translation quality, cultural nuance, and provide detailed explanations of model failures.

Talent poolRecently verified$25-40/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Dutch

This role involves evaluating and correcting AI-generated Dutch text, focusing on natural language flow, appropriate formality levels, and regional variations. Dutch experts will assess translations and model outputs, rewriting substandard content to match native speaker standards.

Talent poolRecently verified$25-40/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Finnish

This role involves evaluating AI model outputs on Finnish translation and localization tasks, identifying errors and quality issues, and providing corrected versions. The position is suited for Finnish language specialists who can assess how well AI handles linguistic and cultural nuances in real-world language work.

Talent poolRecently verified$25-40/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: German

This role involves evaluating and correcting German language output from AI models to improve naturalness and accuracy. You will assess translations, rate model outputs, and rewrite text to sound native while teaching AI systems how German is actually spoken, including regional variations and proper formality levels.

Talent poolRecently verified$25-40/hr

FAQ

Questions about AI training jobs

What kinds of jobs are listed here?

AI training, model evaluation, RLHF, red teaming, data labeling and expert review roles, in fields from software and data science to medicine, law, finance, languages and science.

Are these jobs remote?

They are done online. Some are limited to residents of certain countries: when the platform publishes that restriction, the listing shows it.

Do I need prior experience in AI?

Usually not. Most roles ask for professional experience in your own field. Each listing states the experience the platform requires.

How is the work paid?

By the hour or by the task, by the platform that hires you. We show the pay only when the platform publishes it. It is not a guarantee of income or hours.

How do I apply?

Apply now opens the official posting on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac, where you complete the application. You need no account here. See how it works.