OpenRecently verified

AI Tutor - Humanities

  • Science & Research
  • Platform: xAI

About this job

This role asks a humanities scholar to judge and improve how an AI model handles questions in linguistics, history, classical studies, literature, philosophy, ethics, and the arts. The work centers on checking the model's answers against primary sources, flagging errors and unsupported claims, and writing reference answers that separate evidence from interpretation and speculation. It suits candidates with deep scholarly training who write precise, carefully calibrated analysis.

What you'll do

  • Evaluate model outputs for factual accuracy, logical coherence, and hidden assumptions.
  • Write reference answers and datasets grounded in primary sources.
  • Steel-man opposing views and mark what is settled, contested, or unknown.
  • Collaborate with engineering teams to design evaluation tasks and rubrics.

Requirements

LLM

    Pay

    Pay not disclosed

    The platform does not publish the pay for this job.

    Location

    The platform has not published which countries are eligible for this job.

    How to apply

    You apply on xAI's careers board: submit the application for the AI Tutor role, then complete the assessments and interviews described in its posting.

    All xAI jobs and how the platform works

    Source

    Official posting: https://job-boards.greenhouse.io/xai/jobs/5227951007

    Last checked on October 8, 2026.

    Similar jobs

    • Business, Consulting & Operations
    • Platform: Alignerr

    Design and author realistic customer support triage tasks that train and evaluate AI agents on real enterprise work, including ticket triage, account configuration checks, and escalation decisions. This role is for experienced support professionals who can create detailed task scenarios with clear scoring rubrics that reflect actual judgment calls made by support engineers.

    • Languages, Translation & Voice
    • Platform: Alignerr
    • Language: Spanish

    This remote contract role involves reviewing transcripts of AI-assisted healthcare phone calls held in Central American Spanish. Labelers judge how the AI confirms patient details, handles health concerns and poor audio, and explains scheduling and next steps. It suits native or near-native Spanish speakers with basic healthcare knowledge who can follow detailed guidelines.

    • Health & Medicine
    • Platform: micro1
    • Location: 58 countries
    • Language: Indonesian

    This role invites a psychologist holding a doctorate and bilingual in Indonesian-English to contribute to the training of next-generation AI systems. You will analyze clinical cases, develop culturally adapted psychological scenarios in both languages, and assess AI-generated content according to ethical and professional standards. Your expertise in psychology and your mastery of the linguistic and cultural nuances of both languages are essential.

    Posted September 29, 2026

    OpenRecently verified$100-200/hr

    Referral link to micro1's job list: search for this role there. Open this exact job

    • Business, Consulting & Operations
    • Platform: micro1
    • Location: 58 countries
    • Level: Intermediate

    This role invites you to mobilize your expertise in interviewing and investigative questioning to contribute to training next-generation AI systems. You will conduct structured elicitation sessions, analyze responses in real time by detecting verbal and behavioral cues, and document your analyses according to precise frameworks. The ideal profile possesses solid experience in professional interviews, intelligence information gathering, or investigative questioning.

    Posted September 18, 2026

    OpenRecently verified$25-65/hr

    Referral link to micro1's job list: search for this role there. Open this exact job

    • Data & Machine Learning
    • Platform: Mercor
    • Location: United States
    • Language: Danish

    You join a red teaming team specialized in adversarial evaluation of conversational AI models. Your role consists of testing AI systems by exploring their vulnerabilities (jailbreaks, prompt injections, biases), generating high-quality data documenting these vulnerabilities, and producing reproducible reports to strengthen model safety. This position is for bilingual English-Danish experts with prior experience in red teaming or related fields (cybersecurity, adversarial ML, socio-technical analysis).

    Posted July 30, 2026

    OpenRecently verified$48-62/hr

    Referral link: we may earn a fee. Apply without it

    • Data & Machine Learning
    • Platform: Mercor
    • Location: United States
    • Language: Dutch

    This role consists of testing the robustness and security of conversational AI models by subjecting them to sophisticated adversarial attacks. You will generate high-quality training data by identifying vulnerabilities, biases, and systemic risks that automated tests do not detect. The ideal profile has prior experience in red teaming, structured adversarial thinking, and the ability to clearly communicate security risks.

    Posted July 30, 2026

    OpenRecently verified$48-62/hr

    Referral link: we may earn a fee. Apply without it

    Keep exploring