1234 open jobs

Remote AI training jobs

Open roles in AI training, model evaluation, RLHF and data labeling for domain experts. Search, filter, then apply directly on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac.

"prompt"×Clear all
1 to 20 of 68 jobs
  • Data & Machine Learning
  • Platform: Alignerr

This role involves testing and evaluating AI chatbot responses across diverse topics by engaging in conversations, identifying issues, and providing structured feedback. The position is designed for individuals with strong critical thinking and communication skills who want to contribute to AI safety and improvement without requiring prior technical or AI experience.

  • Software Engineering
  • Platform: Alignerr

An experienced penetration tester will conduct offensive security assessments against AI-powered applications, infrastructure, and machine learning pipelines to identify vulnerabilities including prompt injection, model manipulation, and data poisoning attacks. The role involves simulating real-world attack scenarios, documenting findings with remediation guidance, and evaluating AI-specific security risks in a fully remote, flexible contract engagement.

  • Data & Machine Learning
  • Platform: Alignerr

An AI Red Team Analyst probes and stress-tests AI systems to identify security vulnerabilities, craft adversarial prompts, and evaluate model safety before bad actors can exploit them. This remote contract role suits security-minded professionals who combine cybersecurity expertise with hands-on experience testing large language models and AI systems.

  • Software Engineering
  • Platform: Alignerr

Penetration testing experts are needed to identify security vulnerabilities in AI-powered applications and infrastructure through offensive security testing. This remote contract role combines traditional penetration testing methodologies with emerging AI-specific attack scenarios such as prompt injection and model manipulation.

  • Data & Machine Learning
  • Platform: Alignerr

This role involves testing and evaluating AI models by designing adversarial prompts and scenarios to identify weaknesses, biases, and unsafe outputs. Red team testers document discovered failure modes and assess their severity to help improve AI safety before deployment.

  • Software Engineering
  • Platform: Alignerr

This role involves conducting security testing and red-teaming exercises on advanced AI models to identify vulnerabilities, design adversarial prompts, and evaluate safety failures. The specialist will document findings and collaborate with engineering teams to strengthen AI system resilience and reliability.

  • Writing, Creative & Design
  • Platform: Alignerr

This role involves analyzing short video clips and writing precise cinematographic descriptions to help train AI video generation models. The position is ideal for film professionals who can articulate visual techniques using accurate cinema terminology and structured prompts.

  • Software Engineering
  • Platform: Alignerr

This role involves conducting penetration tests and security assessments on AI-powered applications, infrastructure, and machine learning pipelines to identify vulnerabilities and attack vectors. You will probe for weaknesses including prompt injection, model manipulation, and data poisoning, then document findings with remediation guidance for a fully remote contract position.

  • Data & Machine Learning
  • Platform: Alignerr

This contract role asks security-minded people to probe AI systems built on the OpenClaw platform by running red-teaming exercises and writing adversarial prompts. The work suits cybersecurity practitioners who enjoy finding weaknesses and who can document their findings clearly for technical and non-technical readers.

  • Writing, Creative & Design
  • Platform: Alignerr

This remote contract role asks you to write original, inventive prompts that test how advanced AI models reason, respond, and handle edge cases. It suits people with strong writing skills and broad curiosity, and no technical or AI background is needed.

  • Business, Consulting & Operations
  • Platform: Alignerr

This role asks experienced revenue and sales operations professionals to design realistic problems that test AI agents on reconciling bookings, CRM, order, and communication data into one defensible answer. The work covers writing task prompts, scoring rubrics, and task environments, then solving and calibrating each task against AI models. It suits practitioners who know enterprise sales systems and can define checkable standards for a correct answer.

  • Business, Consulting & Operations
  • Platform: Alignerr

This remote hourly contract asks experienced revenue operations and sales operations professionals to build realistic tasks that test AI agents on bookings, account reconciliation and fulfillment work. The author writes prompts and scoring rubrics, sets up task environments, then solves and calibrates each task against AI models. It suits people with hands-on enterprise B2B sales or revenue operations experience.

  • Writing, Creative & Design
  • Platform: Alignerr

This remote contract role involves watching short video clips and turning what appears on screen into structured written prompts covering framing, mood, setting, and narrative. The prompts are used to help AI models learn to understand and recreate visual storytelling. It suits writers with a strong visual sense who can describe moving images clearly without relying on dialogue or audio.

  • Finance & Accounting
  • Platform: DataAnnotation

This role involves evaluating and improving AI models' accounting capabilities by grading their responses against real-world accounting standards and practices. Accountants will write test prompts, review AI outputs for errors in bookkeeping and tax scenarios, and provide correct answers with detailed explanations to help train the models.

Talent poolRecently verified$40-125/hr
  • Science & Research
  • Platform: DataAnnotation

This role involves evaluating AI models' biological reasoning against established scientific literature. You will write test prompts to probe biological understanding, review AI-generated responses for errors and fabricated citations, and provide scientifically accurate answers based on primary sources when the model's output is inadequate.

Talent poolRecently verified$40-125/hr
  • Health & Medicine
  • Platform: DataAnnotation

This role involves evaluating how well AI models reason through cardiology cases, identifying errors in clinical judgment that could harm patients. Cardiologists will write prompts to test AI reasoning, review AI outputs for accuracy against current guidelines, and provide correct answers when the model errs, serving as the training signal for AI improvement.

Talent poolRecently verified$40-125/hr
  • Science & Research
  • Platform: DataAnnotation

This role involves evaluating AI-generated chemical reasoning and responses to identify flawed mechanisms, unsafe procedures, and analytical errors. You will write prompts to test chemical knowledge, review AI outputs for errors, and provide correct expert-level answers with proper attention to safety and experimental conditions.

Talent poolRecently verified$40-125/hr
  • Legal
  • Platform: DataAnnotation

This role involves evaluating AI-generated contract drafting and redlining to identify logical errors, ambiguities, and one-sided terms that automated systems miss. Contract professionals will draft test prompts, review model outputs for issues like inconsistent definitions and broken cross-references, and rewrite clauses to ensure precise legal language.

Talent poolRecently verified$40-125/hr
  • Finance & Accounting
  • Platform: DataAnnotation

A corporate accountant evaluates AI model outputs on financial accounting tasks, checking for errors in GAAP treatment, consolidations, and month-end closing procedures. You will draft test prompts, review AI-generated entries for mistakes, and provide correct treatments with clear working papers for training purposes.

Talent poolRecently verified$40-125/hr
  • Legal
  • Platform: DataAnnotation

This role involves evaluating AI-generated legal work on mergers, acquisitions, financings, and corporate governance to identify logical gaps, unworkable terms, and misaligned risk allocation. You will draft test prompts for AI models handling complex transactional agreements and rewrite provisions where the AI output falls short, applying the discipline of carefully negotiated deals.

Talent poolRecently verified$40-125/hr

FAQ

Questions about AI training jobs

What kinds of jobs are listed here?

AI training, model evaluation, RLHF, red teaming, data labeling and expert review roles, in fields from software and data science to medicine, law, finance, languages and science.

Are these jobs remote?

They are done online. Some are limited to residents of certain countries: when the platform publishes that restriction, the listing shows it.

Do I need prior experience in AI?

Usually not. Most roles ask for professional experience in your own field. Each listing states the experience the platform requires.

How is the work paid?

By the hour or by the task, by the platform that hires you. We show the pay only when the platform publishes it. It is not a guarantee of income or hours.

How do I apply?

Apply now opens the official posting on Mercor, micro1, Turing, Ethos, Mindrift, DataAnnotation, Alignerr, Handshake AI, Surge AI, xAI, Vetto or Terac, where you complete the application. You need no account here. See how it works.