← All jobs
Turing Verified
remote ·

LLM Annotator - Master's Degree

Pay on listing
Share
Data Analysis remote
Posted Sep 14, 2026

About Turing:

Based in San Francisco, California, Turing is the world's leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports customers in two ways: first, by accelerating frontier research with high-quality data, advanced training pipelines, plus top AI researchers who specialize in coding, reasoning, STEM, multilinguality, multimodality, and agents; and second, by applying that expertise to help enterprises transform AI from proof of concept into proprietary intelligence with systems that perform reliably, deliver measurable impact, and drive lasting results on the P&L.


Role Overview

We are seeking highly motivated LLM Annotators to support the evaluation and improvement of cutting-edge Large Language Models (LLMs). In this role, you will analyze structured data, create challenging prompts, evaluate AI-generated responses for factual accuracy and reasoning quality, and provide evidence-backed feedback to improve model performance.


We are looking for curious, detail-oriented professionals who enjoy solving complex problems and evaluating AI systems. The ideal candidate can analyze data, think critically, validate responses against evidence, and clearly articulate why a model's output is correct or incorrect. Experience working with AI models and creating challenging evaluation prompts is a strong advantage.

Key Responsibilities

  • Create challenging prompts that evaluate an LLM's ability to retrieve, analyze, and reason over structured data.
  • Assess AI-generated responses for factual accuracy, logical reasoning, and completeness.
  • Identify model failures, inconsistencies, hallucinations, and reasoning gaps.
  • Validate model outputs using provided datasets and supporting evidence.
  • Document findings with clear, evidence-based explanations.
  • Consistently follow annotation guidelines and maintain high-quality standards.

Minimum Qualifications

  • Master's degree or higher in any discipline.
  • Minimum 3 years of professional, research, or teaching experience.
  • Strong analytical and critical thinking skills.
  • Excellent written English communication skills.
  • Exceptional attention to detail and ability to validate information against source data.

    Preferred Qualifications
  • Experience working with Large Language Models (LLMs) or Generative AI.
  • Familiarity with prompt engineering, AI evaluation, data annotation, or model testing.
  • Experience working with structured datasets (CSV, Excel, databases, etc.).
  • Ability to identify edge cases and design prompts that expose model limitations.

    Benefits
  • Opportunity to work on cutting-edge AI projects.
  • Competitive compensation.
  • Flexible working hours and remote work environment.

Offer Details

  • Commitments Required: 40, 30 or 20 hours per week with at least 4 hours PST overlap
  • Employment type: Contractor assignment (no medical/paid leave)
  • Duration of contract: 4 weeks


    Evaluation Process:
  • Shortlisting based on qualifications and assessment scores.
Pay range
Pay on listing
Share
Similar roles

You might also like

micro1 Verified New
remote · hourly
Big Data Engineer
Data Analysis
Posted Sep 15, 2026
$30–$80/hr
Mercor Verified New
remote · hourly
Application Users - Excel on macOS - Professionals
Data Analysis
Posted Sep 13, 2026
$60–$70/hr 40h/wk
Mercor Verified New
remote · hourly
Application Users - EViews on Windows - STEM
Data Analysis
Posted Sep 13, 2026
$45–$55/hr 40h/wk
Mercor Verified New
remote · hourly
Application Users - Stata SE on Windows - STEM
Data Analysis
Posted Sep 13, 2026
$45–$55/hr 40h/wk