hiringremote.Get job alerts
Still open

LLM Annotator - Master's Degree

Turing · Remote

Pay
See listing
Where
Worldwide
Degree
See listing
Posted
22d ago
Checked
today

An older role: first posted 22d ago, and Turing still listed it when we checked today. Newer roles tend to fill faster. See the jobs hiring now.

Before you applyHow to pass the Turing assessment and interviewsMost people who don't get in fail the screening, not the CV. Five minutes here first.

What the work is

Role Overview

We are seeking highly motivated LLM Annotators to support the evaluation and improvement of cutting-edge Large Language Models (LLMs). In this role, you will analyze structured data, create challenging prompts, evaluate AI-generated responses for factual accuracy and reasoning quality, and provide evidence-backed feedback to improve model performance.

We are looking for curious, detail-oriented professionals who enjoy solving complex problems and evaluating AI systems. The ideal candidate can analyze data, think critically, validate responses against evidence, and clearly articulate why a model's output is correct or incorrect. Experience working with AI models and creating challenging evaluation prompts is a strong advantage.

Key Responsibilities

  • Create challenging prompts that evaluate an LLM's ability to retrieve, analyze, and reason over structured data.
  • Assess AI-generated responses for factual accuracy, logical reasoning, and completeness.
  • Identify model failures, inconsistencies, hallucinations, and reasoning gaps.
  • Validate model outputs using provided datasets and supporting evidence.
  • Document findings with clear, evidence-based explanations.
  • Consistently follow annotation guidelines and maintain high-quality standards.

Minimum Qualifications

  • Master's degree or higher in any discipline.
  • Minimum 3 years of professional, research, or teaching experience.
  • Strong analytical and critical thinking skills.
  • Excellent written English communication skills.
  • Exceptional attention to detail and ability to validate information against source data.

Preferred Qualifications

  • Experience working with Large Language Models (LLMs) or Generative AI.
  • Familiarity with prompt engineering, AI evaluation, data annotation, or model testing.
  • Experience working with structured datasets (CSV, Excel, databases, etc.).
  • Ability to identify edge cases and design prompts that expose model limitations.

Benefits

  • Opportunity to work on cutting-edge AI projects.
  • Competitive compensation.
  • Flexible working hours and remote work environment.

Offer Details

  • Commitments Required: 40, 30 or 20 hours per week with at least 4 hours PST overlap
  • Employment type: Contractor assignment (no medical/paid leave)
  • Duration of contract: 4 weeks

Evaluation Process

  • Shortlisting based on qualifications and assessment scores.

Pay

See listing, fully remote. How payouts and tax work.

Check history

How we score →
today–Open
Apply on TuringSee listing