Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • Assess psychological case studies and AI-generated content for accuracy and ethics, and write culturally sensitive scenarios in Lithuanian and English for model training. A psychology PhD with clinical and academic experience is the stated profile, plus native or near-native Lithuanian. 70 openings, remote contractor role, $100–200/hour.

    $100 – $200 / HourWorldwide
  • Full-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior board-certified physicians write golden solutions, instruction specs and clinical benchmarks for frontier models. Hybrid in the Bay Area, 40 hours a week for an initial 6 months, 4+ years post-residency and a US licence. $70–110/hour.

    $70 – $110 / HourHybridOpen to United States
  • Interpret dermatology cases from images and history, annotate lesions to a schema, grade AI assessments and write the criteria they are judged by. Non-clinical, shared expert pool, open worldwide. Board-certified or board-eligible, 3+ years post-residency, 15 hours a week minimum. A flat $270/hour.

    $270 / HourWorldwide
  • Write and review evaluation tasks built on journal manuscripts, abstracts, posters and decks, and judge AI-drafted sections against ICMJE, GPP and EQUATOR guidelines and the source CSRs and TFLs. Five years of publication authoring at a sponsor, CRO or MedComms agency. Fifty openings, contractor, $50–80/hour.

    $50 – $80 / HourWorldwide
  • Residency-trained physicians in any specialty write grading criteria, evaluate multi-turn clinical dialogues and annotate clinical reasoning for AI systems. Shared pool with several workstreams. US only, 3+ years post-residency, 20 hours a week minimum, $150/hour.

    $150 / HourOpen to United States
  • Practising US hospitalists and inpatient internists review H&Ps, progress notes and discharge summaries, and judge AI-written inpatient documentation against what a hospitalist would chart. Needs C1 or better in one of 25 listed languages. 2+ years post-residency, 10 hours a week, a flat $170/hour.

    $170 / HourOpen to United States
  • Paid pilot for a biotech and pharma research team: create and critique rubrics that assess commercial drugs and development programs, judge investment-style theses, and assess 5–10 companies end to end. For specialist-fund biotech analysts with 5+ years. About 10–20 hours over 1–2 weeks, $120–200/hour.

    $120 – $200 / HourWorldwide
  • A full-time research post at micro1: design and own the benchmarks, clinical quality rubrics and validation protocols that measure AI medical reasoning and decision support, and research where models fail. Advanced healthcare degree (MD, DO, PhD, MPH, PharmD) and research experience required. Base salary $200,000–250,000 plus equity.

    $200000 – $250000 / YearWorldwide
  • Annotate healthcare regulatory documents, analyse HIPAA, Stark Law and Anti-Kickback questions, refine compliant contract language and flag legal risk in scenarios for AI training. A US JD, US bar admission and healthcare practice at a BigLaw or top healthcare firm preferred. 10 openings at $140–400/hour.

    $140 – $400 / HourWorldwide
  • Write and review evaluation tasks from DSURs, PSURs/PBRERs and case-level safety data: check ICH E2F and E2C(R2) structure, reconcile case counts across sections, and judge signal and benefit-risk conclusions, to train AI safety-writing tools. Five years in PV or drug safety. Forty openings, contractor, $70–80/hour.

    $70 – $80 / HourWorldwide
  • Paid pilot for board-certified, practising physicians with deep therapeutic-area expertise: interpret trial endpoints, judge whether results would change prescribing and real-world uptake, and write rubrics that evaluate AI analysis of drugs. 10–20 hours over 1–2 weeks. $150–230/hour.

    $150 – $230 / HourWorldwide
  • Evaluate AI-generated psychological content for clinical accuracy, bias and cultural fit, and build case studies and therapeutic dialogues for training. PhD in psychology or MD with psychiatry residency, plus native fluency in one of 16 listed languages and advanced English. 70 openings, contractor, $100–200/hour.

    $100 – $200 / HourWorldwide
  • Full-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior drug development scientists write golden solutions, instruction specs and benchmarks for pharma R&D reasoning. Hybrid in the Bay Area, 40 hours a week for an initial 6 months. PhD, PharmD or MD with 4+ years in industry R&D. $75–115/hour.

    $75 – $115 / HourHybridOpen to United States
  • Hands-on preclinical scientists from companies that develop their own drugs annotate R&D data and model outputs, and advise an AI lab building foundation models for drug discovery. Any modality, 5+ years industry R&D, about 10 hours a week. $60–100/hour.

    $60 – $100 / HourWorldwide
  • Preclinical scientists with direct antibody-drug conjugate or bispecific antibody experience annotate R&D data and advise an AI lab building drug discovery foundation models. In-house asset developers only, 5+ years industry R&D, about 10 hours a week. $60–100/hour.

    $60 – $100 / HourWorldwide
  • Nuclear medicine physicists, radiopharmacy staff and cyclotron RSOs red-team frontier AI models: write benign, dual-use and adversarial prompts from medical isotope work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Paid pilot for US pharma forecasters: build or critique drug launch curves and judge whether forecast assumptions (analogs, ramp, peak share, loss of exclusivity) hold up, while writing rubrics that evaluate AI analysis of drugs. 5+ years, 10–20 hours over 1–2 weeks. $130–210/hour.

    $130 – $210 / HourOpen to United States
  • Board-certified radiologists label findings, write reference reports, grade AI-generated reads and author rubrics for medical imaging AI. Non-clinical, remote worldwide, hourly contract at $200–400/hour with a 15-hour weekly minimum. Weekly pay via Stripe or Wise.

    $200 – $400 / HourWorldwide
  • Paid pilot for US market access and pricing professionals: set and pressure-test gross-to-net and formulary-tier assumptions by drug class, judge whether coverage and rebating assumptions match real payer behaviour, and write rubrics for AI drug analysis. 5+ years, 10–20 hours. $175–200/hour.

    $175 – $200 / HourWorldwide
  • Paid pilot for US epidemiologists: size patient populations for drugs and indications (prevalence, incidence, diagnosed to treated to addressable), judge whether estimates are sound and well sourced, and write rubrics for AI drug analysis. 5+ years and a graduate degree, 10–20 hours. $150–175/hour.

    $150 – $175 / HourWorldwide
  • Write or verify hard ten-option multiple-choice questions for an AI benchmark across clinical medicine, imaging, pharmacovigilance, health economics and rehabilitation, with step-by-step solutions and references. MD, DO, PhD or doctoral candidate, 10+ hours a week, asynchronous. $94–119/hour.

    $94 – $119 / HourWorldwide
  • Clinicians whose main practice is young abuse survivors review AI mental-health guidance on abuse cases, build case scenarios, and write trauma-informed best-practice content so models respond safely to vulnerable youth. MD or PhD (psychology or social work) preferred. $100–200/hour, 8 openings, contractor, remote.

    $100 – $200 / HourWorldwide
  • Write and grade evaluation tasks built on real health authority requests and sponsor responses, judging whether a drafted answer would survive a regulator. Five years of clinical submissions experience is the stated bar. Thirty openings, contractor, $50–80/hour, remote.

    $50 – $80 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.