OpenRecently verified

Senior Machine Learning Engineer

  • Data & Machine Learning
  • Platform: Alignerr

$60-80/hrAs published by the platform.

About this job

This contract role asks engineers to write and review step-by-step reasoning traces that show how a language model should plan, use tools, and reach decisions. The output becomes training data meant to make AI systems more reliable on complex, real-world tasks. It suits experienced machine learning engineers and researchers who can work remotely and on their own schedule.

What you'll do

  • Author detailed reasoning traces covering planning, tool use, and multi-step decisions
  • Design data strategies that help models handle intricate real-world scenarios
  • Review and evaluate reasoning traces for logical consistency, structure, and completeness
  • Break advanced ML problems into documented reasoning steps for model training

Requirements

Machine LearningLLM

    Pay

    $60-80/hr

    Pay as published by the platform. It is not a guarantee of income or hours.

    Location

    The platform has not published which countries are eligible for this job.

    How to apply

    You apply on Alignerr, Labelbox's expert network: create a profile, complete an AI-led interview and skills assessment, then get matched to projects in your field.

    All Alignerr jobs and how the platform works

    Source

    Official posting: https://www.alignerr.com/jobs/0462e105-5e34-4bc1-aa98-6a24a27c8191

    Last checked on October 8, 2026.

    Similar jobs

    • Data & Machine Learning
    • Platform: Vetto

    This role is a hands-on research position focused on post-training large language models, where expert annotations are converted into training data and experimental reward signals. It suits a researcher with a PhD or equivalent experience in machine learning or a related quantitative field who can work independently and iterate fast on messy experiments.

    Posted January 23, 2026

    OpenRecently verifiedPay not disclosed
    • Data & Machine Learning
    • Platform: Alignerr

    A fully remote contract role where PhD-level biochemists create expert-level problems, author gold-standard solutions, and audit AI-generated biochemical reasoning to train and improve scientific AI models. Ideal for biochemists passionate about scientific rigor who want to shape the accuracy of next-generation AI without requiring prior AI experience.

    • Data & Machine Learning
    • Platform: Alignerr

    This role involves evaluating and improving AI models that process visual data, including images and video. Computer vision and machine learning experts will assess model outputs, identify failure modes, create training annotations, and provide technical feedback to guide the development of AI systems used at scale.

    OpenRecently verified$100-150/hr
    • Data & Machine Learning
    • Platform: Handshake AI
    • Location: United States
    • Level: Entry level

    Handshake AI seeks skilled PCB and EDA tool users to evaluate AI-generated content and provide feedback on electronics design tasks. Contributors will work flexibly on project-based contract work that involves assessing large language model responses related to PCB layout, schematic capture, and embedded hardware design.

    Posted March 10, 2025

    OpenRecently verifiedUp to $125/hr
    • Data & Machine Learning
    • Platform: DataAnnotation
    • Level: Intermediate

    This role involves evaluating how well AI models can perform computational biology tasks. Computational biologists and bioinformaticians will design realistic analysis scenarios from their own work, run them through AI systems, and grade the results against professional standards to help benchmark model capabilities.

    Talent poolRecently verified$40-125/hr
    • Data & Machine Learning
    • Platform: DataAnnotation

    As a Machine Learning Engineer, you will evaluate and improve how AI models reason about machine learning systems, including training dynamics, evaluation design, and deployment strategies. You will write prompts to test model reasoning, identify subtle errors in AI outputs, and provide correct solutions based on real practitioner experience.

    Talent poolRecently verified$40-150/hr

    Keep exploring