OpenRecently verified

AI Security Tester

  • Data & Machine Learning
  • Platform: Alignerr

$15-75/hrAs published by the platform.

About this job

This role involves conducting security testing and red-teaming exercises on AI systems to identify vulnerabilities, weaknesses, and failure modes. The position is designed for security professionals who can think like attackers and use creative methods to stress-test AI models and evaluate their safety, bias, and policy compliance.

What you'll do

  • Conduct red-teaming exercises to uncover security weaknesses and failure modes in AI systems.
  • Design and execute adversarial prompts, jailbreak attempts, and edge-case scenarios to test model guardrails.
  • Evaluate AI outputs for safety, bias, harmful content, and policy compliance.
  • Document vulnerabilities, unexpected behaviors, and exploits in clear, structured reports.

Requirements

LLM

    Pay

    $15-75/hr

    Pay as published by the platform. It is not a guarantee of income or hours.

    Location

    The platform has not published which countries are eligible for this job.

    How to apply

    You apply on Alignerr, Labelbox's expert network: create a profile, complete an AI-led interview and skills assessment, then get matched to projects in your field.

    All Alignerr jobs and how the platform works

    Source

    Official posting: https://www.alignerr.com/jobs/098bd9f7-eeb4-46cd-8123-1cc8be651091

    Last checked on October 8, 2026.

    Similar jobs

    • Data & Machine Learning
    • Platform: Alignerr

    An AI Red Team Analyst probes and stress-tests AI systems to identify security vulnerabilities, craft adversarial prompts, and evaluate model safety before bad actors can exploit them. This remote contract role suits security-minded professionals who combine cybersecurity expertise with hands-on experience testing large language models and AI systems.

    • Data & Machine Learning
    • Platform: Alignerr

    This contract role asks security-minded people to probe AI systems built on the OpenClaw platform by running red-teaming exercises and writing adversarial prompts. The work suits cybersecurity practitioners who enjoy finding weaknesses and who can document their findings clearly for technical and non-technical readers.

    • Data & Machine Learning
    • Platform: Mercor
    • Location: United States
    • Language: Danish

    You join a red teaming team specialized in adversarial evaluation of conversational AI models. Your role consists of testing AI systems by exploring their vulnerabilities (jailbreaks, prompt injections, biases), generating high-quality data documenting these vulnerabilities, and producing reproducible reports to strengthen model safety. This position is for bilingual English-Danish experts with prior experience in red teaming or related fields (cybersecurity, adversarial ML, socio-technical analysis).

    Posted July 30, 2026

    OpenRecently verified$48-62/hr

    Referral link: we may earn a fee. Apply without it

    • Data & Machine Learning
    • Platform: Mercor
    • Location: United States
    • Language: Dutch

    This role consists of testing the robustness and security of conversational AI models by subjecting them to sophisticated adversarial attacks. You will generate high-quality training data by identifying vulnerabilities, biases, and systemic risks that automated tests do not detect. The ideal profile has prior experience in red teaming, structured adversarial thinking, and the ability to clearly communicate security risks.

    Posted July 30, 2026

    OpenRecently verified$48-62/hr

    Referral link: we may earn a fee. Apply without it

    • Software Engineering
    • Platform: DataAnnotation

    This role involves evaluating how AI models reason about offensive security concepts and identifying flaws in their exploit chain reasoning. Penetration testers with hands-on red team experience will test model outputs across reconnaissance, exploitation, privilege escalation, and lateral movement, then write accurate attack paths that reflect real-world tradecraft when models fall short.

    Talent poolRecently verified$40-125/hr
    • Data & Machine Learning
    • Platform: Terac
    • Location: United States

    This is a paid one-hour trial for people who know a professional or personal workflow well. Participants turn that workflow into a demanding prompt, test it in ChatGPT to find where the model fails, and make the prompt harder when the model succeeds. The work suits domain experts and power users who can judge AI outputs quickly and who are open to ongoing evaluation work.

    Posted August 25, 2026

    OpenRecently verifiedPay not disclosed

    Keep exploring