1234 open jobs

Remote AI training jobs

Open roles in AI training, model evaluation, RLHF and data labeling for domain experts. Search, filter, then apply directly on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac.

DataAnnotation×Clear all
21 to 40 of 113 jobs
  • Finance & Accounting
  • Platform: DataAnnotation

A Credit Analyst evaluates AI-generated credit analyses to identify hidden default risks and validate underwriting decisions. You will assess cash-flow coverage, leverage ratios, covenants, and collateral while writing credit decisions that serve as training data for machine learning models.

Talent poolRecently verified$40-125/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Croatian

This role involves judging how well frontier AI models handle Croatian translation, localization, and cultural nuance on real language projects. The contributor finds errors, explains why they occur, and writes better answers when a model falls short. It is aimed at linguists and bilingual reviewers with professional language experience who want flexible, remote, project-based work.

Talent poolRecently verified$25-40/hr
  • Software Engineering
  • Platform: DataAnnotation

This role involves evaluating AI models' ability to generate secure code by identifying vulnerabilities, testing their security reasoning, and writing corrections when they fail. It suits security engineers and cybersecurity professionals who can spot weaknesses in authentication, cryptography, input handling, and secret management that automated tools miss.

Talent poolRecently verified$40-125/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Czech

This role involves evaluating AI model outputs on Czech translation and localization tasks, identifying errors, and providing corrected versions to improve AI training. The position is suited for Czech language specialists who can assess translation quality, cultural nuance, and provide detailed explanations of model failures.

Talent poolRecently verified$25-40/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Danish

This role involves evaluating AI model performance on Danish translation and localization tasks. You will assess the quality of AI-generated content, identify errors, and provide corrections to train better language models. The work is remote and flexible, suited for professionals with Danish language expertise.

Talent poolRecently verified$25-40/hr
  • Data & Machine Learning
  • Platform: DataAnnotation

As a Data Scientist, you will evaluate analyses generated by AI models on real datasets, checking whether the statistical methods and conclusions are sound. You will identify flaws in the reasoning, stress-test the logic, and write detailed assessments that serve as training data for future model improvements.

Talent poolRecently verified$40-150/hr
  • Writing, Creative & Design
  • Platform: DataAnnotation

Evaluate AI-generated design work to identify usability flaws, accessibility problems, and poor design decisions that may be hidden beneath polished aesthetics. Provide detailed critiques that explain what good design reasoning looks like, helping train AI models to reason about user experience and visual hierarchy.

Talent poolRecently verified$40-125/hr
  • Software Engineering
  • Platform: DataAnnotation

Evaluate how AI models handle DevOps tasks such as CI/CD pipelines, infrastructure-as-code, and cloud operations. Identify failures in automation, security configurations, and deployment processes, then write correct solutions that a platform engineer would trust.

Talent poolRecently verified$40-150/hr
  • Science & Research
  • Platform: DataAnnotation
  • Level: Intermediate

A drug discovery scientist will design realistic preclinical tasks and evaluate how well frontier AI models perform them, grading responses against professional standards. The role involves creating scenarios based on real workflows, running them through AI systems, and providing detailed feedback on whether the models correctly conduct SAR analysis, DMPK interpretation, screening triage, and other core discovery work.

Talent poolRecently verified$40-125/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Dutch

This role involves evaluating and correcting AI-generated Dutch text, focusing on natural language flow, appropriate formality levels, and regional variations. Dutch experts will assess translations and model outputs, rewriting substandard content to match native speaker standards.

Talent poolRecently verified$25-40/hr
  • Engineering
  • Platform: DataAnnotation
  • Level: Intermediate

This role involves evaluating how AI models tackle real-world electrical and electronics engineering problems by designing challenging circuit and systems problems, running them through frontier AI models, and judging the model outputs against professional engineering standards. The position is suited for experienced electrical engineers who can provide detailed technical critiques and written rationales to help train AI systems to reason correctly about circuits, systems, and physical constraints.

Talent poolRecently verified$40-125/hr
  • Engineering
  • Platform: DataAnnotation
  • Level: Intermediate

This role involves evaluating how advanced AI models handle electronics and electrical engineering problems. You will design realistic engineering challenges, test them against frontier AI systems, and judge the output quality using your professional expertise, providing detailed critiques and reasoning to help train better AI tools.

Talent poolRecently verified$40-125/hr
  • Engineering
  • Platform: DataAnnotation
  • Level: Intermediate

DataAnnotation seeks an experienced embedded systems developer to evaluate how AI models handle real electrical and electronics engineering problems. You will design realistic engineering challenges, run them through frontier AI models, judge the output against professional standards, and provide detailed written critiques that help train the next generation of AI engineering tools.

Talent poolRecently verified$40-125/hr
  • Finance & Accounting
  • Platform: DataAnnotation

This role involves evaluating and correcting AI-generated equity research notes on public companies, ensuring that financial models, valuations, and investment theses are internally consistent and meet professional standards. Candidates will review estimates and write-ups, reconcile discrepancies between models and narratives, and produce publishable research notes with sell-side rigor.

Talent poolRecently verified$40-125/hr
  • Science & Research
  • Platform: DataAnnotation
  • Level: Intermediate

This role involves designing realistic experimental biology tasks and evaluating how frontier AI models perform on benchmark scientific challenges. You'll create scenarios from your own laboratory practice, run them through AI systems, and grade the AI's output against professional standards to help train models on hands-on experimental work.

Talent poolRecently verified$40-125/hr
  • Finance & Accounting
  • Platform: DataAnnotation

This role involves evaluating financial models built by AI systems on real corporate scenarios, checking the validity of valuation assumptions and accounting mechanics, and writing authoritative analyses that demonstrate sound finance practices. You will stress-test three-statement models, DCF analyses, and working capital logic to identify errors and provide corrections that meet professional standards.

Talent poolRecently verified$40-125/hr
  • Finance & Accounting
  • Platform: DataAnnotation

As a Financial Reporting Manager, you will evaluate AI-generated financial reporting and technical accounting work for accuracy under US GAAP standards, identifying errors in revenue recognition, lease accounting, business combinations, and SEC disclosures. You will review AI outputs for misapplied standards and missing disclosures, then write correct memos and documentation with proper citations when needed.

Talent poolRecently verified$40-125/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: Finnish

This role involves evaluating AI model outputs on Finnish translation and localization tasks, identifying errors and quality issues, and providing corrected versions. The position is suited for Finnish language specialists who can assess how well AI handles linguistic and cultural nuances in real-world language work.

Talent poolRecently verified$25-40/hr
  • Finance & Accounting
  • Platform: DataAnnotation

Evaluate financial planning and forecasting models created by AI systems, identify logical flaws and unrealistic assumptions, and write corrected analyses that finance teams can present to business stakeholders. This role is for experienced FP&A professionals who can distinguish between superficially polished models and those with sound underlying logic.

Talent poolRecently verified$40-125/hr
  • Languages, Translation & Voice
  • Platform: DataAnnotation
  • Language: French

This role involves evaluating and correcting French language output from AI models, focusing on fluency, formality, regional accuracy, and native idiom. Qualified candidates will have native or near-native French proficiency and will provide feedback to improve how models handle grammar, register, and natural speech patterns.

Talent poolRecently verified$25-40/hr

FAQ

Questions about AI training jobs

What kinds of jobs are listed here?

AI training, model evaluation, RLHF, red teaming, data labeling and expert review roles, in fields from software and data science to medicine, law, finance, languages and science.

Are these jobs remote?

They are done online. Some are limited to residents of certain countries: when the platform publishes that restriction, the listing shows it.

Do I need prior experience in AI?

Usually not. Most roles ask for professional experience in your own field. Each listing states the experience the platform requires.

How is the work paid?

By the hour or by the task, by the platform that hires you. We show the pay only when the platform publishes it. It is not a guarantee of income or hours.

How do I apply?

Apply now opens the official posting on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac, where you complete the application. You need no account here. See how it works.