Talent poolRecently verified

Software Engineer

  • Software Engineering
  • Platform: DataAnnotation

$40-150/hrAs published by the platform.

About this job

You will evaluate AI-generated code by running models on real engineering tasks, analyzing their output against production standards, and stress-testing them to find failures. This role is ideal for experienced software engineers who want to contribute to AI model improvement through rigorous code review and red-teaming.

What you'll do

  • Run coding sessions with frontier models on real tasks such as bug fixes, API extensions, and feature implementations.
  • Analyze AI-generated code against production standards to identify bugs and explain failure modes.
  • Red-team models to expose unsafe code, fake test passes, and incorrect solutions.
  • Document failures and provide explanations that serve as training signals for future models.

Requirements

PythonJavaScriptRustC++

    Pay

    $40-150/hr

    Pay as published by the platform. It is not a guarantee of income or hours.

    Location

    The platform has not published which countries are eligible for this job.

    How to apply

    You apply on DataAnnotation: create an account, take the qualification assessment for your field, then pick up paid projects when you qualify.

    All DataAnnotation jobs and how the platform works

    Source

    Official posting: https://www.dataannotation.tech/job-board/software-engineer

    Last checked on October 8, 2026.

    Similar jobs

    • Software Engineering
    • Platform: Alignerr
    • Level: Intermediate

    This remote contract role asks senior full-stack engineers to assess AI-generated code across frontend and backend, judging correctness, quality, and security. It also involves building internal tooling for data annotation and reviewing system designs. It suits experienced engineers who can give precise written feedback that helps AI models write better code.

    • Software Engineering
    • Platform: micro1
    • Location: 58 countries

    This role involves contributing to an AI training project as an open source software engineer. You will explore unfamiliar repositories, implement and debug front-end and back-end components, and evaluate patches generated by AI. The ideal candidate is an experienced developer with significant contributions to established open source projects.

    Posted September 16, 2026

    OpenRecently verified$100-150/hr

    Referral link to micro1's job list: search for this role there. Open this exact job

    • Software Engineering
    • Platform: Alignerr
    • Level: Intermediate

    This remote contract role asks experienced full-stack developers to review and evaluate code and system designs produced by AI models. The work also includes building tooling for data annotation and quality control, done asynchronously on a flexible schedule. It suits engineers with several years of production experience who can give precise technical feedback.

    • Software Engineering
    • Platform: micro1
    • Location: 58 countries

    You will contribute to training next-generation AI models by creating realistic coding tasks and implementing deterministic verifiers. This flexible contracting role is aimed at experienced backend engineers capable of generating relevant technical challenges and automatically evaluating solutions.

    Posted October 6, 2026

    OpenRecently verified$30-100/hr

    Referral link to micro1's job list: search for this role there. Open this exact job

    • Software Engineering
    • Platform: Terac
    • Location: United States

    This remote study asks practicing software engineers to review programming tasks and the evaluation harnesses built to test AI agents on them. You will check whether the tasks, test cases and environment structure reflect realistic software engineering work. It suits engineers with solid experience building, testing and reviewing complex systems.

    Posted October 2, 2026

    • Software Engineering
    • Platform: Terac
    • Location: Argentina

    This paid study asks experienced software engineers to review proposed programming tasks and the harnesses that test AI agents, checking their logic, structure and realism. The work takes place in a screen-shared session that mixes code review, technical discussion and direct feedback on task design. It is aimed at engineers who have already built, reviewed or tested evaluation harnesses or coding tasks.

    Posted October 1, 2026

    Keep exploring