Board-certified radiologists label findings, write reference reports, grade AI-generated reads and author rubrics for medical imaging AI. Non-clinical, remote worldwide, hourly contract at $200–400/hour with a 15-hour weekly minimum. Weekly pay via Stripe or Wise.
Remote AI training and data labeling jobs
Filter jobs
Location: Worldwide
- Worldwide238 jobs, applied. Activate to remove
- United States63 jobs
- United Kingdom9 jobs
- Canada5 jobs
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
- English230 jobs
- Germanno roles alongside your other filters
- Spanishno roles alongside your other filters
- French1 job
- Japanese3 jobs
- Portugueseno roles alongside your other filters
Field
- Languages & Linguisticsno roles alongside your other filters
- Audio & Voiceno roles alongside your other filters
- Engineering49 jobs
- Business & Finance30 jobs
- Software & IT32 jobs
- Health & Medicine54 jobs
- Law, Policy & Security35 jobs
- General & Data Collection1 job
- Science & Math41 jobs
- AI Safety & Evaluation20 jobs
- Video, Image & Design3 jobs
- Data, AI & ML13 jobs
- Writing & Education2 jobs
- Other fields5 jobs
Newest
238 open roles matching these filters · page 8 of 10
- $200 – $400 / HourWorldwide
Paid pilot for US market access and pricing professionals: set and pressure-test gross-to-net and formulary-tier assumptions by drug class, judge whether coverage and rebating assumptions match real payer behaviour, and write rubrics for AI drug analysis. 5+ years, 10–20 hours. $175–200/hour.
$175 – $200 / HourWorldwideGrade AI-generated slides, spreadsheets and documents for real-world data science quality, flagging factual, visual and presentation errors in structured written feedback. Needs 5+ years at a top firm in the US, UK, Canada, Australia or New Zealand. $100–150/hour.
$100 – $150 / HourWorldwidePaid pilot for US epidemiologists: size patient populations for drugs and indications (prevalence, incidence, diagnosed to treated to addressable), judge whether estimates are sound and well sourced, and write rubrics for AI drug analysis. 5+ years and a graduate degree, 10–20 hours. $150–175/hour.
$150 – $175 / HourWorldwideAn expert-interview listing for engineers who have shipped production search, especially agentic search in the LLM era: a 25-minute conversational interview about relevance, evaluation and real trade-offs, with a possible paid 30-minute follow-up call at $200. No coding, no take-home. Listed at $80–150 per task.
$80 – $150 / TaskWorldwideGrade AI-generated slides, spreadsheets and documents for real-world software engineering quality, flagging factual, visual and presentation errors with written feedback. Needs 5+ years at a top firm in the US, UK, Canada, Australia or New Zealand. $100–150/hour.
$100 – $150 / HourWorldwideWrite original, research-frontier questions in your own discipline that current AI models cannot answer, with sourced and cited reference answers, then test and harden them. Open to any field. Paid per accepted task, $40–90/hour band, five openings, remote contractor.
$40 – $90 / HourWorldwideUmbrella listing for Mercor's energetic materials red-team panel: chemists and engineers or operators write benign, dual-use and adversarial prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 16 hired this month.
$65 – $75 / TaskWorldwideUmbrella listing for Mercor's radiological safety red-team panel: RSOs, health physicists, source security, emergency response and nuclear medicine specialists write prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 8 hired this month.
$65 – $75 / TaskWorldwideRF, microwave, antenna and electromagnetics engineers solve and critique hard technical problems for a short-term expert evaluation project: analysing systems and trade-offs, checking calculations and assumptions, and writing rigorous explanations. Hands-on industry experience and an EE-family degree. Remote hourly contract at $80–95/hour.
$80 – $95 / HourWorldwideQA AI-agent runs inside the Financial Forecaster planning app: check the agent used the right scenario, account, coordinate and basis, catch plausible-but-wrong answers, harden tasks and sharpen grading. Needs weekly hands-on Financial Forecaster use and 5+ years in FP&A, reporting, technical accounting or lender reporting. $70–110/hour, remote.
$70 – $110 / HourWorldwideWrite or verify hard ten-option multiple-choice questions for an AI benchmark across clinical medicine, imaging, pharmacovigilance, health economics and rehabilitation, with step-by-step solutions and references. MD, DO, PhD or doctoral candidate, 10+ hours a week, asynchronous. $94–119/hour.
$94 – $119 / HourWorldwideWrite or verify 10-option multiple-choice benchmark questions in applied maths (signal processing, actuarial science, optimization, climate modeling and more), with chain-of-thought solutions and references. For maths PhDs and doctoral candidates. Remote, 10+ hours a week, $61–77/hour.
$61 – $77 / HourWorldwideAuthor executable scientific-computing problems in ecology, biochemistry and genetics for Sci Code, a new AI benchmark: source a paper, dataset or repo, write the prompt and grading criteria, and keep it only if frontier models mostly fail. PhD plus Python or R, Git and Docker. 6 weeks, 20+ hours a week, $70/hour.
$70 / HourWorldwideSenior materials scientists, and electrical or mechanical engineers, author realistic tasks with a prompt, a data room and a grading method, run them against an AI model and tighten them until the model can no longer reason through cleanly. Daily onboarding and office hours. Remote hourly contract at $60–90/hour.
$60 – $90 / HourWorldwideGrade AI-generated slides, spreadsheets and documents for real-world finance quality, catching factual, visual and presentation errors and writing structured feedback. Needs 5+ years at a top firm in the US, UK, Canada, Australia or New Zealand. $100–150/hour.
$100 – $150 / HourWorldwideA full-time, salaried research role at micro1 designing benchmarks, rubrics, datasets and evaluation pipelines for frontier coding agents. Base salary $200,000–260,000 plus equity and benefits, remote, one opening. Three years in software engineering, ML or evaluation.
$200000 – $260000 / YearWorldwideClinicians whose main practice is young abuse survivors review AI mental-health guidance on abuse cases, build case scenarios, and write trauma-informed best-practice content so models respond safely to vulnerable youth. MD or PhD (psychology or social work) preferred. $100–200/hour, 8 openings, contractor, remote.
$100 – $200 / HourWorldwideA part-time fellowship for federal civil litigators: draft and evaluate motions, judge where AI-written advocacy falls short of persuasive, and build the grading criteria that measure it. Requires at least three documented federal motion wins. 50 openings, $150–300/hour, remote contractor.
$150 – $300 / HourWorldwideEngineers with problem-setting credentials (olympiad committees, licensure exam items, qualifying exams) write hard technical problems in aerodynamics, materials, CAD, circuit design or mechanical engineering, and grade AI answers to them. 47 openings, paid per accepted task, $60–110/hour band.
$60 – $110 / HourWorldwideSolve, audit or adjudicate research-level benchmark problems in quantum optics: cascaded optical parametric amplifiers, SU(1,1) interferometers, loss and two-mode squeezing, using Bogoliubov and covariance-matrix methods. Ten openings, contractor, $80–160/hour, fully remote.
$80 – $160 / HourWorldwideWrite and grade evaluation tasks built on real health authority requests and sponsor responses, judging whether a drafted answer would survive a regulator. Five years of clinical submissions experience is the stated bar. Thirty openings, contractor, $50–80/hour, remote.
$50 – $80 / HourWorldwideA full-time research engineering role at micro1 building RL environments, reward functions, verifiers, synthetic data pipelines and automated evaluation systems. Base salary $200,000–300,000 plus equity and benefits, remote, one opening. Deep reinforcement learning experience required.
$200000 – $300000 / YearWorldwideRedline contracts in simulated negotiations, grade AI responses to contract scenarios, and write the grading criteria. Despite the IP title, the stated bar is a US JD, active US bar admission and three years on technology transactions (MSAs, NDAs, DPAs). 100 openings, remote contractor work at $100–200/hour.
$100 – $200 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.