OpenRecently verified

AI Red Team Tester

  • Data & Machine Learning
  • Platform: Alignerr

$15-120/hrAs published by the platform.

About this job

This role involves testing and evaluating AI models by designing adversarial prompts and scenarios to identify weaknesses, biases, and unsafe outputs. Red team testers document discovered failure modes and assess their severity to help improve AI safety before deployment.

What you'll do

  • Design inventive prompts and conversational strategies to probe AI model weaknesses and elicit problematic outputs
  • Document discovered failure modes with clear, reproducible steps and assess the severity of each issue
  • Evaluate AI systems across a variety of models and task types to identify incorrect, unsafe, or biased responses
  • Collaborate asynchronously with AI safety and research teams to share findings and contribute to model improvements

Requirements

Adversarial thinkingProblem documentationCritical analysisStructured writing

    Pay

    $15-120/hr

    Pay as published by the platform. It is not a guarantee of income or hours.

    Location

    The platform has not published which countries are eligible for this job.

    How to apply

    You apply on Alignerr, Labelbox's expert network: create a profile, complete an AI-led interview and skills assessment, then get matched to projects in your field.

    All Alignerr jobs and how the platform works

    Source

    Official posting: https://www.alignerr.com/jobs/04bd31fe-3b6b-431f-9297-aa10922ce6c7

    Last checked on October 8, 2026.

    Similar jobs

    • Data & Machine Learning
    • Platform: Alignerr

    This role involves testing and evaluating AI chatbot responses across diverse topics by engaging in conversations, identifying issues, and providing structured feedback. The position is designed for individuals with strong critical thinking and communication skills who want to contribute to AI safety and improvement without requiring prior technical or AI experience.

    • Data & Machine Learning
    • Platform: Alignerr

    This contract role asks security-minded people to probe AI systems built on the OpenClaw platform by running red-teaming exercises and writing adversarial prompts. The work suits cybersecurity practitioners who enjoy finding weaknesses and who can document their findings clearly for technical and non-technical readers.

    • Data & Machine Learning
    • Platform: Mercor
    • Location: United States
    • Language: Turkish

    You will participate in improving AI model safety by evaluating how they handle sensitive subjects in Turkish. Your linguistic and cultural judgments will help identify and strengthen weaknesses in these systems when facing delicate content. No prior AI experience is required.

    Posted September 4, 2026

    OpenRecently verified$23-27/hr

    Referral link: we may earn a fee. Apply without it

    • Data & Machine Learning
    • Platform: Terac
    • Location: United States

    This is a paid one-hour trial for people who know a professional or personal workflow well. Participants turn that workflow into a demanding prompt, test it in ChatGPT to find where the model fails, and make the prompt harder when the model succeeds. The work suits domain experts and power users who can judge AI outputs quickly and who are open to ongoing evaluation work.

    Posted August 25, 2026

    OpenRecently verifiedPay not disclosed
    • Data & Machine Learning
    • Platform: Mercor
    • Location: United States
    • Language: Danish

    You join a red teaming team specialized in adversarial evaluation of conversational AI models. Your role consists of testing AI systems by exploring their vulnerabilities (jailbreaks, prompt injections, biases), generating high-quality data documenting these vulnerabilities, and producing reproducible reports to strengthen model safety. This position is for bilingual English-Danish experts with prior experience in red teaming or related fields (cybersecurity, adversarial ML, socio-technical analysis).

    Posted July 30, 2026

    OpenRecently verified$48-62/hr

    Referral link: we may earn a fee. Apply without it

    • Software Engineering
    • Platform: DataAnnotation

    This role involves evaluating how AI models reason about offensive security concepts and identifying flaws in their exploit chain reasoning. Penetration testers with hands-on red team experience will test model outputs across reconnaissance, exploitation, privilege escalation, and lateral movement, then write accurate attack paths that reflect real-world tradecraft when models fall short.

    Talent poolRecently verified$40-125/hr

    Keep exploring