1234 open jobs

Remote AI training jobs

Open roles in AI training, model evaluation, RLHF and data labeling for domain experts. Search, filter, then apply directly on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac.

"evaluation"×Clear all
1 to 20 of 319 jobs
  • Finance & Accounting
  • Platform: Alignerr
  • Level: Intermediate

This role involves evaluating and improving AI systems trained on accounting and financial reporting content. Experienced accounting instructors will review AI-generated accounting questions, explanations, and scenarios to assess accuracy and identify errors, providing structured feedback to enhance the quality of AI financial tools.

  • Data & Machine Learning
  • Platform: Alignerr

This role involves testing and evaluating AI chatbot responses across diverse topics by engaging in conversations, identifying issues, and providing structured feedback. The position is designed for individuals with strong critical thinking and communication skills who want to contribute to AI safety and improvement without requiring prior technical or AI experience.

  • Data & Machine Learning
  • Platform: Alignerr

Review and evaluate AI-generated outputs across text, images, and structured data to assess quality, accuracy, and consistency. This remote contract role is ideal for detail-oriented individuals who can identify errors and provide actionable feedback to guide AI model improvement, with no prior AI experience required.

  • Data & Machine Learning
  • Platform: Alignerr

This role involves reviewing AI-generated content to identify and flag bias, safety issues, and ethical concerns that could affect global audiences. Reviewers apply fairness and safety guidelines to ensure AI systems are responsible and trustworthy, providing structured written feedback on diverse content types.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Chinese

This role involves evaluating and improving AI systems designed for Chinese language learning by assessing the linguistic accuracy, naturalness, and educational quality of AI-generated speech and text, as well as learner output across proficiency levels. The position is ideal for Chinese language experts with teaching or evaluation experience who want to directly influence how AI understands and teaches Chinese.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: French

Evaluate and improve AI systems trained on French-language content by assessing the linguistic accuracy, naturalness, and educational quality of AI-generated French speech and text. This role is for French language experts with teaching backgrounds who will provide structured feedback to enhance AI tutoring and language-learning models for learners at various proficiency levels.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Italian

This role involves evaluating and improving AI systems trained on Italian-language content by assessing the linguistic accuracy and naturalness of AI-generated speech and text, as well as evaluating learner language across proficiency levels. The position is ideal for Italian language experts with teaching experience who want to shape how AI understands and teaches Italian.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Japanese

This role involves evaluating and improving AI systems trained on Japanese-language content by assessing the linguistic accuracy, naturalness, and educational quality of AI-generated speech and text. Native Japanese speakers with teaching or language evaluation experience will provide expert feedback to enhance AI tutoring and language-learning models.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Korean

Evaluate and improve Korean language AI systems by assessing AI-generated speech and text, as well as learner language across proficiency levels. This role combines linguistic expertise with educational knowledge to refine AI tutoring and language-learning models through detailed, structured feedback.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Portuguese

This role involves evaluating and improving AI systems trained on Portuguese by assessing AI-generated speech and text, as well as learner language across proficiency levels. Native Portuguese speakers with teaching or assessment experience will provide linguistic expertise and pedagogical insights to enhance AI language understanding and educational quality.

  • Business, Consulting & Operations
  • Platform: Alignerr

This role involves evaluating AI-generated content against safety, fairness, and ethical standards to ensure AI systems align with human values. The position suits those with strong critical thinking and clear communication skills who are interested in ethics and policy, with no prior AI or policy experience required.

  • Data & Machine Learning
  • Platform: Alignerr

This role involves completing structured evaluation and labeling tasks to train and improve AI systems. You will assess AI-generated content across multiple formats and provide feedback to help make AI systems more accurate and reliable, working remotely on a flexible schedule.

  • Data & Machine Learning
  • Platform: Alignerr

AI Training Specialists label, classify, rate, and evaluate AI-generated content across multiple formats to improve AI system accuracy and reliability. This remote, flexible contract role suits detail-oriented individuals committed to quality work on structured AI training tasks.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: French

This role involves recording high-quality French voice samples and evaluating AI-generated speech for naturalness and expressiveness. The contractor will provide feedback to refine AI voice outputs and review scripts for clarity, supporting the development of frontier AI models.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Italian

This role involves recording high-quality voice samples in Italian and evaluating AI-generated speech outputs for naturalness and expressiveness. Contractors will provide feedback to refine AI voice synthesis and review scripts for clarity, working flexibly on an asynchronous schedule to help develop frontier AI speech technology.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Arabic

This role involves recording high-quality Saudi Arabic voice samples and evaluating AI-generated speech to help train and improve AI voice models. The position is ideal for native or near-native Saudi Arabic speakers with voice acting or narration experience who can provide constructive feedback on pronunciation, tone, and naturalness.

  • Finance & Accounting
  • Platform: Alignerr

This role involves evaluating algorithmic trading strategies built into AI systems, stress-testing their logic, and identifying edge cases to ensure sound reasoning in automated trading decisions. It is designed for experienced algorithmic traders who can critically assess trading systems without requiring prior AI training experience.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Arabic

Arabic language experts evaluate and improve AI systems trained on Arabic content by reviewing translations and generated text for accuracy, fluency, and cultural appropriateness. The role is suited for native or near-native Arabic speakers from any professional background who want to influence how AI understands and communicates in Arabic.

  • Languages, Translation & Voice
  • Platform: Alignerr
  • Language: Arabic

This role involves evaluating and improving AI systems trained on Arabic-language content. Arabic language experts will review AI-generated text for linguistic accuracy, cultural appropriateness, and natural phrasing, providing structured feedback to enhance how AI communicates with Arabic speakers worldwide.

  • Languages, Translation & Voice
  • Platform: Alignerr

This role involves evaluating audio recordings for quality, clarity, and accuracy to support AI training. Reviewers will assess recordings for technical issues like noise and distortion, verify transcription accuracy, and provide detailed feedback using evaluation tools, working independently on a flexible schedule.

FAQ

Questions about AI training jobs

What kinds of jobs are listed here?

AI training, model evaluation, RLHF, red teaming, data labeling and expert review roles, in fields from software and data science to medicine, law, finance, languages and science.

Are these jobs remote?

They are done online. Some are limited to residents of certain countries: when the platform publishes that restriction, the listing shows it.

Do I need prior experience in AI?

Usually not. Most roles ask for professional experience in your own field. Each listing states the experience the platform requires.

How is the work paid?

By the hour or by the task, by the platform that hires you. We show the pay only when the platform publishes it. It is not a guarantee of income or hours.

How do I apply?

Apply now opens the official posting on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac, where you complete the application. You need no account here. See how it works.