OpenRecently verified

AI Safety Practitioner

  • Data & Machine Learning
  • Platform: Mercor
  • Location: 40 countries
  • Level: Expert

$60-70/hrAs published by the platform.

About this job

This role involves judging how frontier AI models respond to sensitive and ambiguous topics, checking answers for safety, accuracy and policy compliance. It suits experienced professionals from fields such as psychology, law, science or public policy who can apply safety rubrics and give structured feedback to model developers.

What you'll do

  • Evaluate AI-generated responses for safety, factual accuracy, and policy compliance.
  • Review content on misinformation, self-harm, violence, cyber, biosecurity, and other sensitive domains.
  • Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.
  • Provide structured feedback to improve model alignment and safety performance.

Requirements

AI safetyTrust & SafetyPublic policyScientific researchSecurity
  • 5+ years of experience
  • Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline

Pay

$60-70/hr

Pay as published by the platform. It is not a guarantee of income or hours.

Location

Open to residents of: United States, Denmark, Estonia, Finland, Iceland, Ireland, Latvia, Lithuania, Norway, Sweden, Austria, Belgium, France, Germany, Liechtenstein, Luxembourg, Monaco, Netherlands, Switzerland, United Kingdom, Albania, Bosnia & Herzegovina, Croatia, Greece, Italy, Kosovo, Malta, North Macedonia, Portugal, San Marino, Serbia, Slovenia, Spain, Bulgaria, Czechia, Hungary, Moldova, Poland, Romania, Slovakia.

How to apply

You apply on work.mercor.com: create a profile, complete a skills assessment, then get matched to projects that fit your background.

All Mercor jobs and how the platform works

Source

Official posting: https://work.mercor.com/jobs/list_AAABn2t1L9IIQXb5jIdMlYOz/ai-safety-practitioner

Last checked on October 8, 2026.

Similar jobs

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Bangla

This role involves testing conversational AI models as an AI safety expert by generating adversarial training data to identify vulnerabilities. You will produce reports and structured datasets that enable clients to improve the robustness and security of their AI systems. The ideal profile masters critical content analysis, methodological rigor, and technical communication.

Posted June 4, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Mercor
  • Location: United States
  • Language: Marathi

This role involves testing conversational AI models as a safety expert by generating adversarial attacks, security bypasses, and misuse scenarios to identify vulnerabilities. You will produce high-quality annotation data and reproducible reports that help clients strengthen the robustness of their AI systems. The ideal profile must possess sound judgment about language and content, be rigorous in identifying subtle anomalies, and capable of clearly communicating observations.

Posted June 3, 2026

OpenRecently verified$16-22/hr

Referral link: we may earn a fee. Apply without it

  • Data & Machine Learning
  • Platform: Alignerr

Review and evaluate AI-generated outputs across text, images, and structured data to assess quality, accuracy, and consistency. This remote contract role is ideal for detail-oriented individuals who can identify errors and provide actionable feedback to guide AI model improvement, with no prior AI experience required.

  • Data & Machine Learning
  • Platform: Alignerr

This role involves completing structured evaluation and labeling tasks to train and improve AI systems. You will assess AI-generated content across multiple formats and provide feedback to help make AI systems more accurate and reliable, working remotely on a flexible schedule.

  • Data & Machine Learning
  • Platform: micro1
  • Location: 58 countries
  • Level: Intermediate

This role involves evaluating and monitoring the quality of datasets intended to train next-generation AI systems. You will examine detailed outputs to verify their accuracy, completeness and compliance with guidelines, applying structured evaluation grids and documenting your analyses. The ideal profile masters data review and possesses solid domain expertise, though AI experience is not required.

Posted September 27, 2026

OpenRecently verified$30-60/hr

Referral link to micro1's job list: search for this role there. Open this exact job

  • Data & Machine Learning
  • Platform: Terac
  • Location: United States

This project asks evaluators to examine how a digital shopping assistant answers real e-commerce questions and to locate where its reasoning or product suggestions fall short. The work is remote and ongoing, and it is aimed at people with quality assurance, data evaluation, or e-commerce backgrounds who can build structured evaluation frameworks.

Posted September 2, 2026

Keep exploring

Apply on Mercor (opens Mercor in a new tab)

This is a referral link: we may earn a fee if you are selected and meet the platform's conditions.

Apply without the referral link