Talent poolRecently verified

Penetration Tester

  • Software Engineering
  • Platform: DataAnnotation

$40-125/hrAs published by the platform.

About this job

This role involves evaluating how AI models reason about offensive security concepts and identifying flaws in their exploit chain reasoning. Penetration testers with hands-on red team experience will test model outputs across reconnaissance, exploitation, privilege escalation, and lateral movement, then write accurate attack paths that reflect real-world tradecraft when models fall short.

What you'll do

  • Test model reasoning across offensive security phases including recon, exploitation, privilege escalation, and lateral movement.
  • Identify and break exploit chains that contain filtered payloads, impossible preconditions, or skipped steps.
  • Write accurate attack paths using real operator tradecraft when model outputs are insufficient or incorrect.
  • Evaluate where models bluff or make confident but incorrect claims about offensive security concepts.

Requirements

Penetration TestingRed TeamingExploit DevelopmentOffensive SecurityPrivilege EscalationLateral MovementAttack Chain AnalysisSecurity Tradecraft

    Pay

    $40-125/hr

    Pay as published by the platform. It is not a guarantee of income or hours.

    Location

    The platform has not published which countries are eligible for this job.

    How to apply

    You apply on DataAnnotation: create an account, take the qualification assessment for your field, then pick up paid projects when you qualify.

    All DataAnnotation jobs and how the platform works

    Source

    Official posting: https://www.dataannotation.tech/job-board/penetration-tester

    Last checked on October 8, 2026.

    Similar jobs

    • Software Engineering
    • Platform: Alignerr

    This role involves analyzing security vulnerabilities and threat scenarios in AI systems and large language models to identify weaknesses and recommend mitigations. Security professionals will conduct adversarial testing, evaluate real-world attack scenarios, and provide structured feedback to improve AI safety and resilience.

    • Software Engineering
    • Platform: Alignerr

    Penetration testing experts are needed to identify security vulnerabilities in AI-powered applications and infrastructure through offensive security testing. This remote contract role combines traditional penetration testing methodologies with emerging AI-specific attack scenarios such as prompt injection and model manipulation.

    • Data & Machine Learning
    • Platform: Mercor
    • Location: United States
    • Language: Danish

    You join a red teaming team specialized in adversarial evaluation of conversational AI models. Your role consists of testing AI systems by exploring their vulnerabilities (jailbreaks, prompt injections, biases), generating high-quality data documenting these vulnerabilities, and producing reproducible reports to strengthen model safety. This position is for bilingual English-Danish experts with prior experience in red teaming or related fields (cybersecurity, adversarial ML, socio-technical analysis).

    Posted July 30, 2026

    OpenRecently verified$48-62/hr

    Referral link: we may earn a fee. Apply without it

    • Data & Machine Learning
    • Platform: Mercor
    • Location: 40 countries
    • Level: Expert

    This role involves stress-testing frontier AI models with adversarial prompts to uncover jailbreaks, unsafe behaviour and policy failures on sensitive topics. It is aimed at experienced safety, security, life sciences or policy professionals who can document weaknesses and help researchers strengthen model alignment.

    Posted July 16, 2026

    OpenRecently verified$70-84/hr

    Referral link: we may earn a fee. Apply without it

    • Software Engineering
    • Platform: DataAnnotation

    This role involves evaluating AI models' ability to generate secure code by identifying vulnerabilities, testing their security reasoning, and writing corrections when they fail. It suits security engineers and cybersecurity professionals who can spot weaknesses in authentication, cryptography, input handling, and secret management that automated tools miss.

    Talent poolRecently verified$40-125/hr
    • Software Engineering
    • Platform: DataAnnotation

    You will evaluate AI-generated code by running models on real engineering tasks, analyzing their output against production standards, and stress-testing them to find failures. This role is ideal for experienced software engineers who want to contribute to AI model improvement through rigorous code review and red-teaming.

    Talent poolRecently verified$40-150/hr

    Keep exploring