Bias Detection AI training jobs

Open AI training and model evaluation roles that call for Bias Detection expertise, gathered from 5 platforms. Part of Data & Machine Learning.

33
open jobs
1
posted this week
$55/hr
median published rate
$200/hr
highest published rate
5
platforms hiring

Pay figures cover the 31 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 8, 2026.

21 to 33 of 33 jobs
  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Dutch

This role consists of testing the robustness and security of conversational AI models by subjecting them to sophisticated adversarial attacks. You will generate high-quality training data by identifying vulnerabilities, biases, and systemic risks that automated tests do not detect. The ideal profile has prior experience in red teaming, structured adversarial thinking, and the ability to clearly communicate security risks.

Posted July 30, 2026

OpenRecently verified$48-62/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Indonesian

You will join an adversarial testing team to evaluate the robustness and safety of conversational AI models. Your job is to identify vulnerabilities by generating malicious inputs, annotating failures, and documenting systemic risks to enable clients to improve their AI systems. This role suits experts with experience in security testing and a natural ability to explore the limits of systems.

Posted July 30, 2026

OpenRecently verified$17-25/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Malay

This role involves training conversational AI models by testing them adversarially to identify vulnerabilities and improve their safety. You will generate high-quality training data by annotating failures, classifying risks, and documenting reproducible attack cases. The ideal profile has prior experience in red teaming or cybersecurity, and masters a structured and methodical approach to testing the limits of AI systems.

Posted July 30, 2026

OpenRecently verified$17-25/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: 40 countries
  • Level: Expert

This role involves stress-testing frontier AI models with adversarial prompts to uncover jailbreaks, unsafe behaviour and policy failures on sensitive topics. It is aimed at experienced safety, security, life sciences or policy professionals who can document weaknesses and help researchers strengthen model alignment.

Posted July 16, 2026

OpenRecently verified$70-84/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Assamese

You join a red teaming team to test the robustness and security of conversational AI models. Your mission is to generate adversarial inputs (jailbreaks, prompt injections, misuse cases) to uncover vulnerabilities in AI systems before their deployment. You annotate failures, classify risks, and produce documented datasets that clients can leverage to improve their models.

Posted June 4, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Bangla

This role involves testing conversational AI models as an AI safety expert by generating adversarial training data to identify vulnerabilities. You will produce reports and structured datasets that enable clients to improve the robustness and security of their AI systems. The ideal profile masters critical content analysis, methodological rigor, and technical communication.

Posted June 4, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Gujarati

You join a red teaming team to test and secure conversational AI models. Your job consists of generating adversarial inputs, identifying vulnerabilities (biases, disinformation, harmful behaviors) and producing high-quality annotated data that makes AI safer. This role requires fluent mastery of English and Gujarati as well as refined judgment on language and content.

Posted June 4, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Kannada

This role involves evaluating the safety of conversational AI models by testing them with adversarial inputs and sophisticated attack techniques. You will generate high-quality annotation data, document discovered vulnerabilities, and contribute to strengthening the robustness of AI systems. This job is suited for someone capable of carefully analyzing AI responses, detecting biases and subtle flaws, and communicating findings in a structured manner.

Posted June 4, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Malayalam

You join a specialized AI safety team to test conversational language models by exposing their flaws and vulnerabilities through adversarial attacks. The job consists of generating high-quality training data through error annotation, risk classification, and documentation of reproducible attacks. This role is for rigorous language experts with critical judgment about the quality and accuracy of AI responses.

Posted June 4, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Punjabi

Join a red teaming team specialized in identifying vulnerabilities of conversational AI models. You will generate high-quality training data by testing AI systems with adversarial inputs, documenting flaws, and producing reproducible reports that clients can use to strengthen the security of their models.

Posted June 4, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Tamil

This role involves testing and evaluating the robustness of conversational AI models by subjecting them to adversarial attacks, bypasses, and malicious use cases. You will generate high-quality annotation data to identify vulnerabilities, biases, and systemic risks, following established taxonomies and benchmarks. This role is suited for AI safety experts with advanced proficiency in both English and Tamil.

Posted June 4, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Telugu

You join an adversarial testing team tasked with detecting vulnerabilities in conversational AI models by generating high-quality human-generated data. Your job consists of exploring security flaws, biases and harmful behaviors through malicious inputs, then documenting systemic risks so that clients can correct them. You must be bilingual in English and Telugu and demonstrate rigor in analyzing AI responses on sensitive topics.

Posted June 4, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Vetto

This remote project asks contributors to role-play users holding assigned political views in multi-turn conversations with an AI system, then judge whether the AI stays neutral, balanced and responsible under pushback. It suits people who already take part in political debate in everyday life and can keep a persona separate from their own opinions while applying detailed written guidelines.

Posted May 22, 2026

OpenRecently verifiedPay not disclosed

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about Bias Detection AI training jobs

How many Bias Detection AI training jobs are open right now?

33 Bias Detection AI training jobs are open on SideHustler today. 1 was posted in the last 7 days. Listings are rechecked against the official postings: the latest check on this list was on October 8, 2026.

How much do Bias Detection AI training jobs pay?

Among the 31 open roles that publish an hourly rate in USD, pay runs from $15 to $200 per hour, with a median of $55. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer Bias Detection AI training jobs?

Mercor (16), micro1 (10), Alignerr (5), Vetto (1) and xAI (1). You apply on the platform itself, which handles screening, contracts and payment.

What skills do Bias Detection AI training jobs ask for most?

The most requested areas of expertise on these listings are AI Evaluation, Bias Detection, Vulnerability Assessment, AI Safety, Red Teaming. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.