Red Teaming AI training jobs

Open AI training and model evaluation roles that call for Red Teaming expertise, gathered from 6 platforms. Part of Data & Machine Learning.

84
open jobs
$69/hr
median published rate
$250/hr
highest published rate
6
platforms hiring

Pay figures cover the 82 open jobs that publish an hourly rate in USD. They are published rates, not a guarantee of income. Last checked on October 8, 2026.

41 to 60 of 84 jobs
  • Engineering
  • Platform: Mercor
  • Location: United States

You will contribute to evaluating frontier AI models as a nuclear fuel cycle expert. Your role will consist of writing nuanced technical questions and judging whether the model's answers comply with a safety policy, distinguishing between legitimate requests and potentially dangerous queries. This work requires a deep understanding of the parameters that differentiate a civilian application from a diversion for proliferation purposes.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

This job involves evaluating the capacity of AI models to distinguish legitimate requests from malicious uses in the field of nuclear medicine and medical isotopes. You will write nuanced test prompts and judge whether the model's responses comply with a security policy, while explaining your decisions in writing. This role requires specialized expertise to draw the line between routine professional questions and those concealing dangerous intentions.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

You will participate in evaluating cutting-edge AI models as a nuclear security specialist. You will write calibrated test prompts across three risk levels to judge whether the model correctly distinguishes legitimate professional questions from dangerous requests, then you will evaluate and document its responses according to a defined policy standard.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Engineering
  • Platform: Mercor
  • Location: United States

You are an engineer specialized in propulsion and pyrotechnic systems. You will write calibrated technical questions to test an AI model's ability to distinguish legitimate requests from bypass attempts, then evaluate its responses and document the technical reasoning that justifies each judgment.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

Join a team of radiation safety experts to test the robustness of AI models against requests related to dual-use applications. You will write nuanced evaluation questions, assess model responses against strict safety criteria, and document your technical judgment to distinguish legitimate requests from dangerous ones.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

Mercor is looking for nuclear and safeguards experts to test the robustness of AI models against dual-use requests. You will write targeted prompts on the knife's edge between legitimate professional questions and dangerous requests, then evaluate the model's responses against a defined policy and write reference answers. This role is for practitioners with concrete experience in nuclear material verification and detecting diversion attempts.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

You will contribute to evaluating frontier AI models as a source security specialist. Your role is to write prompts that test the model's ability to distinguish legitimate professional questions from potentially dangerous requests, then evaluate its responses against a defined policy standard. This work requires practical expertise in Category 1 and 2 radioactive source security, as well as strong technical writing skills to justify your assessments.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

You will contribute to evaluating AI models as a specialist chemist in energetic materials, by writing technical requests that test the model's ability to distinguish legitimate questions from dangerous requests. You will need to master the fine line between civilian applications and misuse to design realistic scenarios and evaluate the model's responses.

Posted September 14, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

You are a chemistry expert with practical experience in synthesis, analytical chemistry, or chemical safety. You participate in evaluating AI models by writing calibrated test prompts across three risk levels, then assessing model responses against an established safety framework. This role requires both advanced technical expertise and the ability to document your judgments in an accessible way.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

This role involves evaluating an AI model's responses to requests involving energetic materials and explosives, distinguishing between legitimate questions and dangerous requests. You will write progressively more complex test prompts, assess the model's responses against a defined safety policy, and provide written justification for your technical judgments. This role is intended for experts in materials chemistry specializing in propulsion and initiation systems, with solid practical experience in the field.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Science & Research
  • Platform: Mercor
  • Location: United States

Mercor is looking for nuclear engineering and nuclear safeguards experts to test the capabilities of frontier AI models to assess diversion risk. You will write complex prompts in your field, evaluate responses against a defined policy, and provide written technical justifications. This role requires deep expertise in fuel cycle or non-proliferation, and demonstrated ability to communicate technical issues with clarity.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Health & Medicine
  • Platform: Mercor
  • Location: United States

Mercor is recruiting radiological safety experts to test the robustness of AI models against dual-use requests. You will write progressively complex prompts, evaluate model responses according to a defined policy, and document reference responses with technical justifications. This role requires deep expertise capable of distinguishing legitimate professional questions from potentially dangerous requests.

Posted September 9, 2026

OpenRecently verified$65-75/hr

Referral link: we may earn a fee. Apply without it

  • Business, Consulting & Operations
  • Platform: Mercor
  • Location: United States
  • Language: Croatian

This role involves evaluating and improving the safety of AI models by analyzing their behavior on sensitive topics in Croatian. You will write expert prompts, classify content according to structured guidelines, and identify adversarial formulations, as an English-Croatian bilingual speaker with no prior AI experience required.

Posted September 4, 2026

OpenRecently verified$38-42/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: Mercor
  • Location: Belgium
  • Language: Dutch

This role asks bilingual English and Dutch speakers to test how AI models respond to sensitive topics in Dutch as used in Belgium. The work combines language fluency with cultural judgment, and no prior AI experience is needed because the workflow is taught on the job.

Posted September 4, 2026

OpenRecently verified$48-52/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: Mercor
  • Location: United States
  • Language: Finnish

This role involves evaluating and improving the safety of advanced AI models in Finnish as a bilingual English-Finnish speaker. You will apply your language mastery and judgment to examine how these models handle sensitive topics, without requiring prior experience in AI or machine learning.

Posted September 4, 2026

OpenRecently verified$48-52/hr

Referral link: we may earn a fee. Apply without it

  • Languages, Translation & Voice
  • Platform: Mercor
  • Location: United States
  • Language: French

This role consists of evaluating the safety of AI models in French by examining how they handle sensitive topics. You will write expert prompts, apply classification guidelines, and identify adversarial formulations, leveraging your bilingual mastery of French and English.

Posted September 4, 2026

OpenRecently verified$48-52/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Norwegian

This role consists of evaluating and strengthening the safety of advanced AI models by analyzing their behavior on sensitive topics in Norwegian. You will use your bilingual proficiency (English-Norwegian) and cultural judgment to identify gaps and bypass attempts, with no prior AI experience required.

Posted September 4, 2026

OpenRecently verified$58-62/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Turkish

You will participate in improving AI model safety by evaluating how they handle sensitive subjects in Turkish. Your linguistic and cultural judgments will help identify and strengthen weaknesses in these systems when facing delicate content. No prior AI experience is required.

Posted September 4, 2026

OpenRecently verified$23-27/hr

Referral link: we may earn a fee. Apply without it

  • Generalist & Data Labeling
  • Platform: Mercor
  • Location: United States
  • Language: Ukrainian

This role involves contributing to AI model safety as a bilingual English-Ukrainian expert. You will evaluate how these models handle sensitive topics in Ukrainian, by writing expert prompts and classifying conversations according to structured guidelines. No prior experience in AI or machine learning is required.

Posted September 4, 2026

OpenRecently verified$38-42/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This is a paid one-hour trial for people who know a professional or personal workflow well. Participants turn that workflow into a demanding prompt, test it in ChatGPT to find where the model fails, and make the prompt harder when the model succeeds. The work suits domain experts and power users who can judge AI outputs quickly and who are open to ongoing evaluation work.

Posted August 25, 2026

OpenRecently verifiedPay not disclosed

See which of these jobs match your experience

Add your resume once. Every open job is scored against your profile, and we email you only when a strong match appears. Free.

Get recommended jobs

FAQ

Questions about Red Teaming AI training jobs

How many Red Teaming AI training jobs are open right now?

84 Red Teaming AI training jobs are open on SideHustler today. Listings are rechecked against the official postings: the latest check on this list was on October 8, 2026.

How much do Red Teaming AI training jobs pay?

Among the 82 open roles that publish an hourly rate in USD, pay runs from $14 to $250 per hour, with a median of $69. These are the rates the platforms publish, not a guarantee of income or hours.

Which platforms offer Red Teaming AI training jobs?

Mercor (56), Alignerr (23), DataAnnotation (2), micro1 (1), Terac (1) and Vetto (1). You apply on the platform itself, which handles screening, contracts and payment.

What skills do Red Teaming AI training jobs ask for most?

The most requested areas of expertise on these listings are Red Teaming, Adversarial Testing, AI Evaluation, Vulnerability Assessment, AI Safety. Each listing states its own requirements.

How do I apply?

Apply now opens the official posting on the platform, where you complete the application. You need no account on SideHustler. To see which roles fit your background first, add your resume and every open job is scored against your profile.