Practising US primary care, family medicine or general internal medicine physicians review outpatient notes and judge AI-written documentation against what a PCP would actually chart. Needs C1 or better in one of 25 listed languages. 2+ years post-residency, 10 hours a week, $170–190/hour.
Remote AI training and data labeling jobs
Filter jobs
Location
- Worldwide62 jobs
- United States16 jobs
- United Kingdomno roles alongside your other filters
- Canadano roles alongside your other filters
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
Field: Health & Medicine
- Languages & Linguistics119 jobs
- Audio & Voice114 jobs
- Engineering109 jobs
- Business & Finance107 jobs
- Software & IT88 jobs
- Health & Medicine78 jobs, applied. Activate to remove
- Law, Policy & Security70 jobs
- General & Data Collection68 jobs
- Science & Math65 jobs
- AI Safety & Evaluation63 jobs
- Video, Image & Design44 jobs
- Data, AI & ML35 jobs
- Writing & Education14 jobs
- Other fields6 jobs
Newest
78 open roles matching these filters · page 3 of 4
- $170 – $190 / HourOpen to United States
Full-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior drug development scientists write golden solutions, instruction specs and benchmarks for pharma R&D reasoning. Hybrid in the Bay Area, 40 hours a week for an initial 6 months. PhD, PharmD or MD with 4+ years in industry R&D. $75–115/hour.
$75 – $115 / HourHybridOpen to United StatesHands-on preclinical scientists from companies that develop their own drugs annotate R&D data and model outputs, and advise an AI lab building foundation models for drug discovery. Any modality, 5+ years industry R&D, about 10 hours a week. $60–100/hour.
$60 – $100 / HourWorldwidePreclinical scientists with direct antibody-drug conjugate or bispecific antibody experience annotate R&D data and advise an AI lab building drug discovery foundation models. In-house asset developers only, 5+ years industry R&D, about 10 hours a week. $60–100/hour.
$60 – $100 / HourWorldwideNuclear medicine physicists, radiopharmacy staff and cyclotron RSOs red-team frontier AI models: write benign, dual-use and adversarial prompts from medical isotope work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwidePediatric acute care RNs evaluate AI outputs built from nursing flowsheet documentation, annotate pediatric assessment data and help write clinical benchmarks. US licence outside California, inpatient bedside work within the past 10 years. $55–65/hour.
$55 – $65 / HourOpen to United StatesInpatient registered nurses with broad clinical exposure support ongoing clinical AI product work: annotation, clinical review, model evaluation, user-feedback investigation and guideline development, alongside engineers and clinicians. US only, 10 hours a week minimum, $55–65/hour.
$55 – $65 / HourOpen to United StatesCertified medical coders (CCDS, CHC, CCS or CPC) review clinical documentation and AI-generated notes for coding integrity, validate code-to-documentation alignment and annotate against guidelines. US-based, C1 in any second language, 10 hours a week. $45–65/hour.
$45 – $65 / HourOpen to United StatesPaid pilot for US pharma forecasters: build or critique drug launch curves and judge whether forecast assumptions (analogs, ramp, peak share, loss of exclusivity) hold up, while writing rubrics that evaluate AI analysis of drugs. 5+ years, 10–20 hours over 1–2 weeks. $130–210/hour.
$130 – $210 / HourOpen to United StatesBoard-certified radiologists label findings, write reference reports, grade AI-generated reads and author rubrics for medical imaging AI. Non-clinical, remote worldwide, hourly contract at $200–400/hour with a 15-hour weekly minimum. Weekly pay via Stripe or Wise.
$200 – $400 / HourWorldwidePaid pilot for US market access and pricing professionals: set and pressure-test gross-to-net and formulary-tier assumptions by drug class, judge whether coverage and rebating assumptions match real payer behaviour, and write rubrics for AI drug analysis. 5+ years, 10–20 hours. $175–200/hour.
$175 – $200 / HourWorldwidePaid pilot for US epidemiologists: size patient populations for drugs and indications (prevalence, incidence, diagnosed to treated to addressable), judge whether estimates are sound and well sourced, and write rubrics for AI drug analysis. 5+ years and a graduate degree, 10–20 hours. $150–175/hour.
$150 – $175 / HourWorldwideCertified pharmacy technicians answer medication questions and review AI responses for an AI lab building prior authorization workflows: dosing, interactions, indications, PA requirements. US only, 30–40 hours a week during the project. A flat $35/hour.
$35 / HourOpen to United StatesWrite or verify hard ten-option multiple-choice questions for an AI benchmark across clinical medicine, imaging, pharmacovigilance, health economics and rehabilitation, with step-by-step solutions and references. MD, DO, PhD or doctoral candidate, 10+ hours a week, asynchronous. $94–119/hour.
$94 – $119 / HourWorldwideClinicians whose main practice is young abuse survivors review AI mental-health guidance on abuse cases, build case scenarios, and write trauma-informed best-practice content so models respond safely to vulnerable youth. MD or PhD (psychology or social work) preferred. $100–200/hour, 8 openings, contractor, remote.
$100 – $200 / HourWorldwideWrite and grade evaluation tasks built on real health authority requests and sponsor responses, judging whether a drafted answer would survive a regulator. Five years of clinical submissions experience is the stated bar. Thirty openings, contractor, $50–80/hour, remote.
$50 – $80 / HourWorldwideBuild realistic biostatistics benchmark tasks (clinical trial analysis, regulatory review, observational studies) with real datasets and 35+ item grading rubrics. Paid per accepted task, quoted at $60–100/hour. MS or PhD plus four years in pharma, CRO, hospital or academic medicine. 50 openings.
$60 – $100 / HourWorldwideCheck inpatient records (FHIR data, HPI, PMH, assessments and plans) against the claims made about them, flag missing or inconsistent information, and apply guidelines consistently to improve AI training data. Board-certified MD or DO with inpatient work in the last two years. Thirty openings, contractor, $70–100/hour.
$70 – $100 / HourWorldwideBring ambulance dispatch experience to an AI training project: triage and route simulated emergency requests across radio, phone and chat using standard protocols, document each decision, and help set the quality bar for realistic emergency-communications data. Certification not required. $30–55/hour, 50 openings.
$30 – $55 / HourWorldwideAudit outpatient professional fee coding, review other coders' work, and make calibrated calls on E/M leveling, modifier -25, time vs MDM and NCCI bundling, at about ten reviews a day, for AI training data. CPC required, CPMA preferred, ten years' coding or academic medical centre audit experience. Fifty openings, $30–40/hour.
$30 – $40 / HourWorldwideWrite hard, original medical Q&A pairs that stump frontier AI, source and cite the answers from primary literature and guidelines, and make questions harder when models get them right. Open to medical students, residents and physicians. Paid per accepted task with a weekly minimum. Five openings, $40–90/hour.
$40 – $90 / HourWorldwideRead clinical images of skin lesions, describe them in precise terms, judge how well AI assessments hold up, and write the criteria that define a high-quality dermatological read. Non-clinical, no patient care. Active US licence and five years post-residency. A flat $270/hour.
$270 / HourOpen to United StatesJudge AI-generated clinical notes against what a practising outpatient physician would actually document, and help write the annotation guidelines. Any specialty, but you need C1 or better in Czech, Catalan, Danish, Dutch, Vietnamese or Finnish. US only, 10 hours a week, $170–190/hour.
$170 – $190 / HourOpen to United StatesAssess adolescent eating disorder cases in writing, applying DSM-5, EDE-Q and SCOFF, judging medical stability against MEED, and saying whether FBT or CBT-E fits the clinical stage. Wants a licensed clinician with a recent adolescent caseload, and accepts paediatric medicine alongside psychiatry and psychology. Remote contractor work.
$45 – $95 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.