Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • Quantum and computational chemistry PhDs write original, runnable research problems for a scientific-computing AI benchmark, with grading criteria, calibrated until frontier models fail them more often than they pass. 6 weeks at 20+ hours a week, Git and Docker workflow. Flat $70/hour; 182 hired this month.

    $70 / HourWorldwide
  • Fuel-cycle and criticality safety engineers red-team frontier AI models: write benign, dual-use and adversarial prompts from enrichment, fabrication, reprocessing or criticality work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Hands-on preclinical scientists from companies that develop their own drugs annotate R&D data and model outputs, and advise an AI lab building foundation models for drug discovery. Any modality, 5+ years industry R&D, about 10 hours a week. $60–100/hour.

    $60 – $100 / HourWorldwide
  • A paid research study for a major tech company building AI design tools: senior product, UX and brand designers take part in recorded remote sessions on how experts judge craft and decide what is ready to ship, either as interviewee or interviewer. About 4–5 hours per session. $150–250/hour, scope still being finalised.

    $150 – $250 / HourWorldwide
  • Evaluate AI-generated legal analyses of everyday civil problems (housing, family, consumer, foreclosure, debt) and write structured feedback on reasoning, procedure and access-to-justice issues. For US-admitted lawyers with pro bono program experience. Remote hourly contract at $170/hour.

    $170 / HourWorldwide
  • Remote hourly contract for consultants, analysts and marketers who build decks in PowerPoint on their own Windows PC. $60–70/hour, paid weekly via Stripe or Wise. Your own Microsoft licence and a screen setup above 2.5 megapixels are required; the ad does not describe the tasks.

    $60 – $70 / HourWorldwide
  • QA engineers, SDETs and test automation engineers review browser-based test workflows for AI-generated web apps, checking that each test is feasible, isolated, deterministic and gives a clean pass or fail before it joins a benchmark dataset. 3+ years and Playwright, Cypress or Selenium experience. Remote hourly contract at $30–60/hour.

    $30 – $60 / HourWorldwide
  • Full-time, on-site-hybrid role in the Bay Area: review AI output on business and sales operations tasks, write instruction specs and golden solutions, and build benchmarks with an AI lab's research team. For senior ops leaders with 4+ years. W-2 via Cincinnatus, $60–100/hour.

    $60 – $100 / HourHybridOpen to United States
  • Current and former IAEA, Euratom and state-system safeguards inspectors and accountancy analysts red-team frontier AI models: write benign, dual-use and adversarial prompts, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Audit Mandarin Chinese speech data for two Amazon Sonic collections: judge annotators' transcripts against audio using fixed error codes, and fix word-level timestamp alignment. Native Mandarin as spoken in mainland China is a hard requirement; all rules and rationales are in English. Remote hourly contract at $21.50/hour.

    $21.50 / HourWorldwide
  • Audit Korean speech data for Amazon's Sonic project: check transcripts against audio with fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native Korean as spoken in South Korea is a hard requirement; rationales are written in English. Remote hourly contract at $25/hour.

    $25 / HourWorldwide
  • Audit Japanese speech data for Amazon's Sonic collections: verify annotators' transcripts against the audio with fixed error codes and pass/fail calls, and correct word-level timestamp alignment. Native Japanese as spoken in Japan is a hard requirement; rationales are in English. Remote hourly contract at $37.50/hour.

    $37.50 / HourWorldwide
  • Test AI chat models in Malayalam and English for safety failures and judge whether their answers are accurate, complete and appropriate. Native Malayalam plus careful evaluation skills; prior red-teaming is a plus rather than the core requirement. Remote hourly contract at $16–22/hour, paid weekly.

    $16 – $22 / HourWorldwide
  • Full-time W-2 coordinator (via Cincinnatus LLC) placed with a leading AI lab's GenAI team, running day-to-day operations on AI training-data projects: turning leadership input into expert guidelines, onboarding experts, tracking quality and deliverables. US-based, 3+ years of coordination plus hands-on AI data experience. $45–55/hour.

    $45 – $55 / HourOpen to United States
  • Preclinical scientists with direct antibody-drug conjugate or bispecific antibody experience annotate R&D data and advise an AI lab building drug discovery foundation models. In-house asset developers only, 5+ years industry R&D, about 10 hours a week. $60–100/hour.

    $60 – $100 / HourWorldwide
  • Nuclear medicine physicists, radiopharmacy staff and cyclotron RSOs red-team frontier AI models: write benign, dual-use and adversarial prompts from medical isotope work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Predict when a named sell-side analyst will publish after a catalyst and what the note will say, reasoning only from a fixed evidence cutoff, then grade AI analyses of the same call. For former lead sell-side analysts or very senior associates in the exact sector, typically 8+ years. $150–250/hour, remote.

    $150 – $250 / HourWorldwide
  • Help run LLM training projects on browsing capabilities: track project and annotator performance in Google Sheets, review annotator quality and handle contributor communications. Ops or project management background preferred, but ownership matters more. Remote hourly contract at $50–60/hour, no location limit published.

    $50 – $60 / HourWorldwide
  • Pediatric acute care RNs evaluate AI outputs built from nursing flowsheet documentation, annotate pediatric assessment data and help write clinical benchmarks. US licence outside California, inpatient bedside work within the past 10 years. $55–65/hour.

    $55 – $65 / HourOpen to United States
  • Belgium-based PhD chemists and biologists who write Belgian Dutch: author specialised science prompts and grade how AI models handle accuracy and dual-use safety. Belgium residence required. Part-time remote at $61–65/hour, $13 above the Belgian Dutch generalist role; 9 hires this month.

    $61 – $65 / HourOpen to Belgium
  • Native Danish speakers write sensitive-topic prompts, classify prompts and conversations, and flag adversarial phrasing to make AI models safer in Danish. Business English and a bachelor's (in progress counts). Part-time remote at $48–52/hour; Denmark or Western Europe preferred, not required.

    $48 – $52 / HourWorldwide
  • Australia-based native English speakers with an authentic Australian accent record one session of about four hours, which a Mercor client uses to clone the voice for its internal customer-experience AI agent. $50–100/hour in USD, no acting experience required. You are licensing your voice, so check the terms before recording.

    $50 – $100 / HourOpen to Australia
  • Read a specific non-G7 central bank's statements, minutes and speeches in the source language, score them dovish to hawkish from a fixed evidence cutoff, and grade AI macro analyses. For former central bank economists or senior local rates and FX strategists, typically 8+ years. $150–250/hour, remote.

    $150 – $250 / HourWorldwide
  • Audit French speech data for Amazon's Sonic project: check annotators' transcripts against audio with fixed error codes and pass/fail calls, and correct word-level timestamp alignment. Native French as spoken in France is a hard requirement; rationales are in English. Joint top rate of the Sonic set at $41.50/hour, remote contract.

    $41.50 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.