1234 open jobs

Remote AI training jobs

Open roles in AI training, model evaluation, RLHF and data labeling for domain experts. Search, filter, then apply directly on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac.

"prompt"×Clear all
41 to 60 of 68 jobs
  • Science & Research
  • Platform: Mercor
  • Location: United States

Mercor is looking for nuclear engineering and nuclear safeguards experts to test the capabilities of frontier AI models to assess diversion risk. You will write complex prompts in your field, evaluate responses against a defined policy, and provide written technical justifications. This role requires deep expertise in fuel cycle or non-proliferation, and demonstrated ability to communicate technical issues with clarity.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

Mercor is recruiting radiological safety experts to test the robustness of AI models against dual-use requests. You will write progressively complex prompts, evaluate model responses according to a defined policy, and document reference responses with technical justifications. This role requires deep expertise capable of distinguishing legitimate professional questions from potentially dangerous requests.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States
  • Language: Croatian

This role involves evaluating and improving the safety of AI models by analyzing their behavior on sensitive topics in Croatian. You will write expert prompts, classify content according to structured guidelines, and identify adversarial formulations, as an English-Croatian bilingual speaker with no prior AI experience required.

Posted September 4, 2026

OpenRecently verified$38-42/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: Mercor
  • Location: United States
  • Language: French

This role consists of evaluating the safety of AI models in French by examining how they handle sensitive topics. You will write expert prompts, apply classification guidelines, and identify adversarial formulations, leveraging your bilingual mastery of French and English.

Posted September 4, 2026

OpenRecently verified$48-52/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States
  • Language: Indonesian

This remote-style role asks PhD scientists who are fluent in both English and Indonesian to help make AI models safer on technical subjects. Contributors write expert prompts in Indonesian, then judge model answers for accuracy, helpfulness, and handling of sensitive material. No prior AI experience is needed because the workflow is taught on the job.

Posted September 4, 2026

OpenRecently verified$19-23/hr

Referral link: we may earn a fee. Apply without it

  • Generalist & Data Labeling
  • Platform: Mercor
  • Location: United States
  • Language: Ukrainian

This role involves contributing to AI model safety as a bilingual English-Ukrainian expert. You will evaluate how these models handle sensitive topics in Ukrainian, by writing expert prompts and classifying conversations according to structured guidelines. No prior experience in AI or machine learning is required.

Posted September 4, 2026

OpenRecently verified$38-42/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This project asks evaluators to examine how a digital shopping assistant answers real e-commerce questions and to locate where its reasoning or product suggestions fall short. The work is remote and ongoing, and it is aimed at people with quality assurance, data evaluation, or e-commerce backgrounds who can build structured evaluation frameworks.

Posted September 2, 2026

  • Finance & Accounting
  • Platform: micro1
  • Location: 58 countries
  • Level: Entry level

You will evaluate and improve the ability of AI models to reason about complex financial questions, particularly in valuation, financial modeling, and real-world investment scenarios. You will create prompts and reference responses based on your operational experience, while flagging errors and reasoning weaknesses in the model.

Posted August 27, 2026

OpenRecently verified$245-280/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Business, Consulting & Operations
  • Platform: micro1
  • Location: 58 countries
  • Level: Entry level

This role consists of evaluating and improving AI model responses to strategy and business consulting problems. You will leverage your consulting firm experience to design quality prompts, write reference responses and provide detailed feedback, applying structured analytical frameworks to identify reasoning gaps.

Posted August 27, 2026

OpenRecently verified$245-280/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This is a paid one-hour trial for people who know a professional or personal workflow well. Participants turn that workflow into a demanding prompt, test it in ChatGPT to find where the model fails, and make the prompt harder when the model succeeds. The work suits domain experts and power users who can judge AI outputs quickly and who are open to ongoing evaluation work.

Posted August 25, 2026

OpenRecently verifiedPay not disclosed
  • Business, Consulting & Operations
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

You leverage your professional experience in corporate environments to train AI systems to better understand and process business documents. You create realistic scenarios involving Excel, PowerPoint and Word, evaluate responses generated by models and provide detailed feedback to improve their capabilities.

Posted August 17, 2026

OpenRecently verified$35-50/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Writing, Creative & Design
  • Platform: Mercor
  • Location: United States
  • Language: Norwegian

You are invited to evaluate the quality of music and lyrics generated by AI models, comparing these creations with detailed standards and published works. This role requires confirmed musical expertise and mastery of Norwegian to analyze the authenticity, creativity, and alignment of productions with the provided prompts.

Posted August 12, 2026

OpenRecently verified$42-78/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Danish

You join a red teaming team specialized in adversarial evaluation of conversational AI models. Your role consists of testing AI systems by exploring their vulnerabilities (jailbreaks, prompt injections, biases), generating high-quality data documenting these vulnerabilities, and producing reproducible reports to strengthen model safety. This position is for bilingual English-Danish experts with prior experience in red teaming or related fields (cybersecurity, adversarial ML, socio-technical analysis).

Posted July 30, 2026

OpenRecently verified$48-62/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Dutch

This role consists of testing the robustness and security of conversational AI models by subjecting them to sophisticated adversarial attacks. You will generate high-quality training data by identifying vulnerabilities, biases, and systemic risks that automated tests do not detect. The ideal profile has prior experience in red teaming, structured adversarial thinking, and the ability to clearly communicate security risks.

Posted July 30, 2026

OpenRecently verified$48-62/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Finnish

This role involves testing the limits and vulnerabilities of conversational AI models using adversarial techniques to identify security risks. You generate high-quality data that enables clients to improve the robustness and reliability of their AI systems. The ideal profile masters red teaming, structured adversarial thinking, and the ability to reproducibly document discovered flaws.

Posted July 30, 2026

OpenRecently verified$48-62/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Indonesian

You will join an adversarial testing team to evaluate the robustness and safety of conversational AI models. Your job is to identify vulnerabilities by generating malicious inputs, annotating failures, and documenting systemic risks to enable clients to improve their AI systems. This role suits experts with experience in security testing and a natural ability to explore the limits of systems.

Posted July 30, 2026

OpenRecently verified$17-25/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Malay

This role involves training conversational AI models by testing them adversarially to identify vulnerabilities and improve their safety. You will generate high-quality training data by annotating failures, classifying risks, and documenting reproducible attack cases. The ideal profile has prior experience in red teaming or cybersecurity, and masters a structured and methodical approach to testing the limits of AI systems.

Posted July 30, 2026

OpenRecently verified$17-25/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Norwegian Bokmål

You join a team dedicated to adversarial testing of conversational AI models, identifying their vulnerabilities to malicious inputs, manipulations, and biases. This role consists of generating high-quality training data to strengthen the robustness and security of AI systems. You are ideally a red teaming expert with prior experience in adversarial security testing and fluent proficiency in English and Norwegian.

Posted July 30, 2026

OpenRecently verified$48-62/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Portuguese

Join a red teaming team specialized in AI safety: you will probe conversational models with adversarial attacks, discover vulnerabilities, and generate data that makes AI systems safer. This role, entirely text-based and team-oriented, will allow you to deploy ethical hacking techniques, risk classification, and documentation so that clients can act on identified flaws.

Posted July 30, 2026

OpenRecently verified$29-45/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Swedish

You join a specialized team focused on adversarial evaluation of conversational AI models. Your role is to identify vulnerabilities in these systems by testing them with malicious inputs, then generate the data and reports that will enable clients to strengthen the security of their products. This position requires fluent mastery of English and Swedish, as well as prior experience in adversarial testing or security.

Posted July 30, 2026

OpenRecently verified$48-62/hr

Referral link: we may earn a fee. Apply without it

FAQ

Questions about AI training jobs

What kinds of jobs are listed here?

AI training, model evaluation, RLHF, red teaming, data labeling and expert review roles, in fields from software and data science to medicine, law, finance, languages and science.

Are these jobs remote?

They are done online. Some are limited to residents of certain countries: when the platform publishes that restriction, the listing shows it.

Do I need prior experience in AI?

Usually not. Most roles ask for professional experience in your own field. Each listing states the experience the platform requires.

How is the work paid?

By the hour or by the task, by the platform that hires you. We show the pay only when the platform publishes it. It is not a guarantee of income or hours.

How do I apply?

Apply now opens the official posting on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac, where you complete the application. You need no account here. See how it works.