OpenRecently verified

Database Systems Evaluation Expert

  • Software Engineering
  • Platform: Mercor
  • Location: United States
  • Level: Expert

$2200/taskAs published by the platform.

About this job

This is a paid evaluation role for database engineers who will design realistic tasks that test how well AI agents handle database work. Each task is a self-contained problem with a reproducible environment, a working reference solution, tests, and grading criteria that separate robust engineering from code that only looks correct. It is aimed at experts with deep database-engine internals experience, not at DBA, SQL analytics, or ETL-only profiles.

What you'll do

  • Design one challenging, self-contained database engineering task.
  • Build checks for correctness, failure handling, concurrency, retries, and regressions.
  • Evaluate performance measures such as latency, throughput, and memory use.
  • Write clear grading criteria for engineering judgment and work with technical reviewers to validate the task.

Requirements

PythonJavaRustC++SQLETL
  • 5+ years of experience
  • PhD in database systems

Pay

$2200/task

Pay as published by the platform. It is not a guarantee of income or hours.

Location

Open to residents of: United States.

How to apply

You apply on work.mercor.com: create a profile, complete a skills assessment, then get matched to projects that fit your background.

All Mercor jobs and how the platform works

Source

Official posting: https://work.mercor.com/jobs/list_AAABoQ9G4DSSNzu_zvdIJYwu/database-systems-evaluation-expert

Last checked on October 8, 2026.

Similar jobs

  • Software Engineering
  • Platform: Alignerr
  • Level: Intermediate

This role involves evaluating and providing feedback on AI-generated Python code for correctness, efficiency, and security while designing complex backend algorithmic solutions. It is designed for experienced backend Python developers who want to contribute to cutting-edge AI research by helping train the next generation of AI systems to write better code.

  • Software Engineering
  • Platform: Alignerr
  • Level: Expert

This role involves reviewing and evaluating AI-generated Rust code to ensure correctness, memory safety, and systems-level performance. Experienced Rust engineers will assess the quality of AI outputs, design challenging problems to test AI understanding, and provide structured feedback to improve how AI models reason about systems programming.

  • Software Engineering
  • Platform: Terac
  • Location: United States

This remote study asks practicing software engineers to review programming tasks and the evaluation harnesses built to test AI agents on them. You will check whether the tasks, test cases and environment structure reflect realistic software engineering work. It suits engineers with solid experience building, testing and reviewing complex systems.

Posted October 2, 2026

  • Software Engineering
  • Platform: micro1
  • Location: 58 countries

This role consists of creating reinforcement learning environments to test the capabilities of AI models in solving complex software engineering problems, including code correction, feature creation, and performance optimization. As an expert software engineer, you will produce quality code examples, debugging strategies, and reproducible reference solutions. The role requires confirmed expertise in programming and an established presence in open source.

Posted September 7, 2026

OpenRecently verified$50-100/hr

Referral link to micro1's job list: search for this role there. Open this exact job

Keep exploring

Apply on Mercor (opens Mercor in a new tab)

This is a referral link: we may earn a fee if you are selected and meet the platform's conditions.

Apply without the referral link