Assess psychological case studies and AI-generated content for accuracy and ethics, and write culturally sensitive scenarios in Lithuanian and English for model training. A psychology PhD with clinical and academic experience is the stated profile, plus native or near-native Lithuanian. 70 openings, remote contractor role, $100–200/hour.
Remote AI training and data labeling jobs
Filter jobs
Location
- Worldwide54 jobs
- United States10 jobs
- United Kingdomno roles alongside your other filters
- Canadano roles alongside your other filters
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
Field: Health & Medicine
- Languages & Linguisticsno roles alongside your other filters
- Audio & Voiceno roles alongside your other filters
- Engineering55 jobs
- Business & Finance47 jobs
- Software & IT42 jobs
- Health & Medicine64 jobs, applied. Activate to remove
- Law, Policy & Security51 jobs
- General & Data Collection2 jobs
- Science & Math47 jobs
- AI Safety & Evaluation22 jobs
- Video, Image & Design6 jobs
- Data, AI & ML18 jobs
- Writing & Education4 jobs
- Other fields5 jobs
Newest
64 open roles matching these filters · page 2 of 3
- $100 – $200 / HourWorldwide
Full-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior board-certified physicians write golden solutions, instruction specs and clinical benchmarks for frontier models. Hybrid in the Bay Area, 40 hours a week for an initial 6 months, 4+ years post-residency and a US licence. $70–110/hour.
$70 – $110 / HourHybridOpen to United StatesInterpret dermatology cases from images and history, annotate lesions to a schema, grade AI assessments and write the criteria they are judged by. Non-clinical, shared expert pool, open worldwide. Board-certified or board-eligible, 3+ years post-residency, 15 hours a week minimum. A flat $270/hour.
$270 / HourWorldwideWrite and review evaluation tasks built on journal manuscripts, abstracts, posters and decks, and judge AI-drafted sections against ICMJE, GPP and EQUATOR guidelines and the source CSRs and TFLs. Five years of publication authoring at a sponsor, CRO or MedComms agency. Fifty openings, contractor, $50–80/hour.
$50 – $80 / HourWorldwideResidency-trained physicians in any specialty write grading criteria, evaluate multi-turn clinical dialogues and annotate clinical reasoning for AI systems. Shared pool with several workstreams. US only, 3+ years post-residency, 20 hours a week minimum, $150/hour.
$150 / HourOpen to United StatesPractising US hospitalists and inpatient internists review H&Ps, progress notes and discharge summaries, and judge AI-written inpatient documentation against what a hospitalist would chart. Needs C1 or better in one of 25 listed languages. 2+ years post-residency, 10 hours a week, a flat $170/hour.
$170 / HourOpen to United StatesPaid pilot for a biotech and pharma research team: create and critique rubrics that assess commercial drugs and development programs, judge investment-style theses, and assess 5–10 companies end to end. For specialist-fund biotech analysts with 5+ years. About 10–20 hours over 1–2 weeks, $120–200/hour.
$120 – $200 / HourWorldwideA full-time research post at micro1: design and own the benchmarks, clinical quality rubrics and validation protocols that measure AI medical reasoning and decision support, and research where models fail. Advanced healthcare degree (MD, DO, PhD, MPH, PharmD) and research experience required. Base salary $200,000–250,000 plus equity.
$200000 – $250000 / YearWorldwideAnnotate healthcare regulatory documents, analyse HIPAA, Stark Law and Anti-Kickback questions, refine compliant contract language and flag legal risk in scenarios for AI training. A US JD, US bar admission and healthcare practice at a BigLaw or top healthcare firm preferred. 10 openings at $140–400/hour.
$140 – $400 / HourWorldwideWrite and review evaluation tasks from DSURs, PSURs/PBRERs and case-level safety data: check ICH E2F and E2C(R2) structure, reconcile case counts across sections, and judge signal and benefit-risk conclusions, to train AI safety-writing tools. Five years in PV or drug safety. Forty openings, contractor, $70–80/hour.
$70 – $80 / HourWorldwidePaid pilot for board-certified, practising physicians with deep therapeutic-area expertise: interpret trial endpoints, judge whether results would change prescribing and real-world uptake, and write rubrics that evaluate AI analysis of drugs. 10–20 hours over 1–2 weeks. $150–230/hour.
$150 – $230 / HourWorldwideEvaluate AI-generated psychological content for clinical accuracy, bias and cultural fit, and build case studies and therapeutic dialogues for training. PhD in psychology or MD with psychiatry residency, plus native fluency in one of 16 listed languages and advanced English. 70 openings, contractor, $100–200/hour.
$100 – $200 / HourWorldwidePractising US primary care, family medicine or general internal medicine physicians review outpatient notes and judge AI-written documentation against what a PCP would actually chart. Needs C1 or better in one of 25 listed languages. 2+ years post-residency, 10 hours a week, $170–190/hour.
$170 – $190 / HourOpen to United StatesFull-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior drug development scientists write golden solutions, instruction specs and benchmarks for pharma R&D reasoning. Hybrid in the Bay Area, 40 hours a week for an initial 6 months. PhD, PharmD or MD with 4+ years in industry R&D. $75–115/hour.
$75 – $115 / HourHybridOpen to United StatesHands-on preclinical scientists from companies that develop their own drugs annotate R&D data and model outputs, and advise an AI lab building foundation models for drug discovery. Any modality, 5+ years industry R&D, about 10 hours a week. $60–100/hour.
$60 – $100 / HourWorldwidePreclinical scientists with direct antibody-drug conjugate or bispecific antibody experience annotate R&D data and advise an AI lab building drug discovery foundation models. In-house asset developers only, 5+ years industry R&D, about 10 hours a week. $60–100/hour.
$60 – $100 / HourWorldwideNuclear medicine physicists, radiopharmacy staff and cyclotron RSOs red-team frontier AI models: write benign, dual-use and adversarial prompts from medical isotope work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwidePaid pilot for US pharma forecasters: build or critique drug launch curves and judge whether forecast assumptions (analogs, ramp, peak share, loss of exclusivity) hold up, while writing rubrics that evaluate AI analysis of drugs. 5+ years, 10–20 hours over 1–2 weeks. $130–210/hour.
$130 – $210 / HourOpen to United StatesBoard-certified radiologists label findings, write reference reports, grade AI-generated reads and author rubrics for medical imaging AI. Non-clinical, remote worldwide, hourly contract at $200–400/hour with a 15-hour weekly minimum. Weekly pay via Stripe or Wise.
$200 – $400 / HourWorldwidePaid pilot for US market access and pricing professionals: set and pressure-test gross-to-net and formulary-tier assumptions by drug class, judge whether coverage and rebating assumptions match real payer behaviour, and write rubrics for AI drug analysis. 5+ years, 10–20 hours. $175–200/hour.
$175 – $200 / HourWorldwidePaid pilot for US epidemiologists: size patient populations for drugs and indications (prevalence, incidence, diagnosed to treated to addressable), judge whether estimates are sound and well sourced, and write rubrics for AI drug analysis. 5+ years and a graduate degree, 10–20 hours. $150–175/hour.
$150 – $175 / HourWorldwideWrite or verify hard ten-option multiple-choice questions for an AI benchmark across clinical medicine, imaging, pharmacovigilance, health economics and rehabilitation, with step-by-step solutions and references. MD, DO, PhD or doctoral candidate, 10+ hours a week, asynchronous. $94–119/hour.
$94 – $119 / HourWorldwideClinicians whose main practice is young abuse survivors review AI mental-health guidance on abuse cases, build case scenarios, and write trauma-informed best-practice content so models respond safely to vulnerable youth. MD or PhD (psychology or social work) preferred. $100–200/hour, 8 openings, contractor, remote.
$100 – $200 / HourWorldwideWrite and grade evaluation tasks built on real health authority requests and sponsor responses, judging whether a drafted answer would survive a regulator. Five years of clinical submissions experience is the stated bar. Thirty openings, contractor, $50–80/hour, remote.
$50 – $80 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.