Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • Red-team frontier AI models on chemical misuse: write benign, dual-use and adversarial prompts from forensic casework, judge the model's answers against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Build AI evaluation tasks in financial reporting, audit and technical accounting: scenarios, reference memos and workpapers, and rubrics that separate senior judgment from exam recall. US GAAP/PCAOB or IFRS/ISA track. 5+ years at a Big Four firm or as a corporate controller. $70–80/hour, remote.

    $70 – $80 / HourWorldwide
  • Read clinical images of skin lesions, describe them in precise terms, judge how well AI assessments hold up, and write the criteria that define a high-quality dermatological read. Non-clinical, no patient care. Active US licence and five years post-residency. A flat $270/hour.

    $270 / HourOpen to United States
  • The UK posting of Mercor's molecular biology project: design primers, plasmids, gRNAs, mRNA constructs and repair templates as ground truth for a frontier model, and write the rubrics that judge them. PhD strongly preferred, first-author record expected, 20 hours a week. $70–105/hour.

    $70 – $105 / HourOpen to United Kingdom
  • Design the primers, plasmids, gRNAs, mRNA constructs and repair templates that become ground truth for a frontier model, and write the rubrics that judge sequence design quality. PhD strongly preferred, first-author record expected, 20 hours a week minimum. US only, $70–105/hour.

    $70 – $105 / HourOpen to United States
  • Read physician and patient research surveys and say whether the clinical terminology, treatment pathways and response options match how the disease is actually treated. A flat $200/hour, three years in a therapeutic area, and an MD is explicitly not required. Fully remote, on your own schedule.

    $200 / HourWorldwide
  • Produce the decks, models and operating-model work you'd build on a normal engagement, alongside experts from two other disciplines, so the outputs can be turned into the tasks and rubrics that train frontier models. $140–200/hour for 40 hours a week from 14 September to 10 October 2026. You must be based in the UK with the right to work there.

    $140 – $200 / HourOpen to United Kingdom
  • Build and document the production pipelines you'd work on during a normal week, alongside experts from two other disciplines, so the outputs can be turned into the tasks and rubrics that train frontier models. $140–200/hour for 40 hours a week from 14 September to 10 October 2026. You must be based in the UK with the right to work there.

    $140 – $200 / HourOpen to United Kingdom
  • Design and review physician and patient questionnaires (screeners, question wording, response scales, branching logic, respondent burden) and say whether an instrument will actually produce usable data. A flat $140/hour, and therapeutic-area specialism is welcome but not required. Fully remote, on your own schedule.

    $140 / HourWorldwide
  • Write difficult problems in your own field for a top AI lab's language models. Open to PhDs (or equivalent industry experience) in medicine, statistics, AI/ML, computer science, game development or mechanical and aerospace engineering, living in the US, UK, Canada or Australia. Remote hourly contract at $55–80/hour, paid weekly.

    $55 – $80 / HourOpen to Australia, Canada and 2 more countries

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.