A short, one-off evaluation task for mental health professionals: compare pairs of simulated clinical conversations and judge which is more realistic. The ad says it takes up to 30 minutes. Remote contract at $100/hour, for clinicians with a relevant degree and practical clinical experience.
Remote AI training and data labeling jobs
Filter jobs
Location
- Worldwide14 jobs
- United States14 jobs
- United Kingdomno roles alongside your other filters
- Canadano roles alongside your other filters
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
- English26 jobs
- German2 jobs
- Spanish2 jobs
- French2 jobs
- Japanese2 jobs
- Portuguese2 jobs
- Dutch3 jobs
- Korean2 jobs
- Chinese2 jobs
- Indonesian2 jobs
- Italian2 jobs
- Danish3 jobs
- Thaino roles alongside your other filters
- Finnish3 jobs
- Russian2 jobs
- Arabic2 jobs
- Vietnamese1 job
- Gujaratino roles alongside your other filters
- Hindino roles alongside your other filters
- Norwegianno roles alongside your other filters
- Tamilno roles alongside your other filters
- Teluguno roles alongside your other filters
- Filipino2 jobs
- Banglano roles alongside your other filters
- Marathino roles alongside your other filters
- Malayno roles alongside your other filters
- Polish2 jobs
- Turkish2 jobs
- Ukrainian2 jobs
- Bulgarian2 jobs
- Catalan3 jobs
- Czech3 jobs
- Hungarian2 jobs
- Georgianno roles alongside your other filters
- Lithuanianno roles alongside your other filters
- Urduno roles alongside your other filters
- Welshno roles alongside your other filters
- Croatianno roles alongside your other filters
- Haitian Creole2 jobs
- Kannadano roles alongside your other filters
- Malayalamno roles alongside your other filters
- Odiano roles alongside your other filters
- Punjabino roles alongside your other filters
- Romanian2 jobs
- Basqueno roles alongside your other filters
- Persianno roles alongside your other filters
- Hebrewno roles alongside your other filters
- Swedishno roles alongside your other filters
- Swahilino roles alongside your other filters
Field: Health & Medicine
- Languages & Linguistics58 jobs
- Audio & Voice43 jobs
- Engineering40 jobs
- Business & Finance52 jobs
- Software & IT32 jobs
- Health & Medicine28 jobs, applied. Activate to remove
- Law, Policy & Security17 jobs
- General & Data Collection42 jobs
- Science & Math41 jobs
- AI Safety & Evaluation61 jobs
- Video, Image & Design9 jobs
- Data, AI & ML15 jobs
- Writing & Education1 job
- Other fieldsno roles alongside your other filters
Newest
28 open roles matching these filters · page 1 of 2
- $100 / HourWorldwide
US-based pharmacists and pharmacy technicians listen to short audio clips of spoken medication names and score whether each is pronounced accurately and clearly, using a rubric. Remote contract at $70/hour for an AI healthcare project, with a calibration exercise first.
$70 / HourOpen to United StatesA paid ~20-minute online survey for current or former Devoted Health Medicare Advantage members aged 64+ in the US, mixing multiple-choice and short voice answers to inform healthcare AI tools. Listed at $120/hour, paid once after verification.
$120 / HourOpen to United StatesFull-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior board-certified physicians write golden solutions, instruction specs and clinical benchmarks for frontier models. Hybrid in the Bay Area, 40 hours a week for an initial 6 months, 4+ years post-residency and a US licence. $70–110/hour.
$70 – $110 / HourHybridOpen to United StatesInterpret dermatology cases from images and history, annotate lesions to a schema, grade AI assessments and write the criteria they are judged by. Non-clinical, shared expert pool, open worldwide. Board-certified or board-eligible, 3+ years post-residency, 15 hours a week minimum. A flat $270/hour.
$270 / HourWorldwideResidency-trained physicians in any specialty write grading criteria, evaluate multi-turn clinical dialogues and annotate clinical reasoning for AI systems. Shared pool with several workstreams. US only, 3+ years post-residency, 20 hours a week minimum, $150/hour.
$150 / HourOpen to United StatesHealthcare back-office specialists (coding, prior authorization, denials and appeals, payer operations) design realistic operational scenarios, build the supporting records, and author and evaluate tasks that train AI agents. 2+ years and currently in role, 10 hours a week. $50–65/hour.
$50 – $65 / HourWorldwidePractising US hospitalists and inpatient internists review H&Ps, progress notes and discharge summaries, and judge AI-written inpatient documentation against what a hospitalist would chart. Needs C1 or better in one of 25 listed languages. 2+ years post-residency, 10 hours a week, a flat $170/hour.
$170 / HourOpen to United StatesPaid pilot for a biotech and pharma research team: create and critique rubrics that assess commercial drugs and development programs, judge investment-style theses, and assess 5–10 companies end to end. For specialist-fund biotech analysts with 5+ years. About 10–20 hours over 1–2 weeks, $120–200/hour.
$120 – $200 / HourWorldwidePaid pilot for board-certified, practising physicians with deep therapeutic-area expertise: interpret trial endpoints, judge whether results would change prescribing and real-world uptake, and write rubrics that evaluate AI analysis of drugs. 10–20 hours over 1–2 weeks. $150–230/hour.
$150 – $230 / HourWorldwidePractising US primary care, family medicine or general internal medicine physicians review outpatient notes and judge AI-written documentation against what a PCP would actually chart. Needs C1 or better in one of 25 listed languages. 2+ years post-residency, 10 hours a week, $170–190/hour.
$170 – $190 / HourOpen to United StatesFull-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior drug development scientists write golden solutions, instruction specs and benchmarks for pharma R&D reasoning. Hybrid in the Bay Area, 40 hours a week for an initial 6 months. PhD, PharmD or MD with 4+ years in industry R&D. $75–115/hour.
$75 – $115 / HourHybridOpen to United StatesHands-on preclinical scientists from companies that develop their own drugs annotate R&D data and model outputs, and advise an AI lab building foundation models for drug discovery. Any modality, 5+ years industry R&D, about 10 hours a week. $60–100/hour.
$60 – $100 / HourWorldwidePreclinical scientists with direct antibody-drug conjugate or bispecific antibody experience annotate R&D data and advise an AI lab building drug discovery foundation models. In-house asset developers only, 5+ years industry R&D, about 10 hours a week. $60–100/hour.
$60 – $100 / HourWorldwideNuclear medicine physicists, radiopharmacy staff and cyclotron RSOs red-team frontier AI models: write benign, dual-use and adversarial prompts from medical isotope work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwidePediatric acute care RNs evaluate AI outputs built from nursing flowsheet documentation, annotate pediatric assessment data and help write clinical benchmarks. US licence outside California, inpatient bedside work within the past 10 years. $55–65/hour.
$55 – $65 / HourOpen to United StatesInpatient registered nurses with broad clinical exposure support ongoing clinical AI product work: annotation, clinical review, model evaluation, user-feedback investigation and guideline development, alongside engineers and clinicians. US only, 10 hours a week minimum, $55–65/hour.
$55 – $65 / HourOpen to United StatesCertified medical coders (CCDS, CHC, CCS or CPC) review clinical documentation and AI-generated notes for coding integrity, validate code-to-documentation alignment and annotate against guidelines. US-based, C1 in any second language, 10 hours a week. $45–65/hour.
$45 – $65 / HourOpen to United StatesPaid pilot for US pharma forecasters: build or critique drug launch curves and judge whether forecast assumptions (analogs, ramp, peak share, loss of exclusivity) hold up, while writing rubrics that evaluate AI analysis of drugs. 5+ years, 10–20 hours over 1–2 weeks. $130–210/hour.
$130 – $210 / HourOpen to United StatesBoard-certified radiologists label findings, write reference reports, grade AI-generated reads and author rubrics for medical imaging AI. Non-clinical, remote worldwide, hourly contract at $200–400/hour with a 15-hour weekly minimum. Weekly pay via Stripe or Wise.
$200 – $400 / HourWorldwidePaid pilot for US market access and pricing professionals: set and pressure-test gross-to-net and formulary-tier assumptions by drug class, judge whether coverage and rebating assumptions match real payer behaviour, and write rubrics for AI drug analysis. 5+ years, 10–20 hours. $175–200/hour.
$175 – $200 / HourWorldwidePaid pilot for US epidemiologists: size patient populations for drugs and indications (prevalence, incidence, diagnosed to treated to addressable), judge whether estimates are sound and well sourced, and write rubrics for AI drug analysis. 5+ years and a graduate degree, 10–20 hours. $150–175/hour.
$150 – $175 / HourWorldwideCertified pharmacy technicians answer medication questions and review AI responses for an AI lab building prior authorization workflows: dosing, interactions, indications, PA requirements. US only, 30–40 hours a week during the project. A flat $35/hour.
$35 / HourOpen to United StatesWrite or verify hard ten-option multiple-choice questions for an AI benchmark across clinical medicine, imaging, pharmacovigilance, health economics and rehabilitation, with step-by-step solutions and references. MD, DO, PhD or doctoral candidate, 10+ hours a week, asynchronous. $94–119/hour.
$94 – $119 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.