Write and review evaluation tasks built on journal manuscripts, abstracts, posters and decks, and judge AI-drafted sections against ICMJE, GPP and EQUATOR guidelines and the source CSRs and TFLs. Five years of publication authoring at a sponsor, CRO or MedComms agency. Fifty openings, contractor, $50–80/hour.
Remote AI training and data labeling jobs
Filter jobs
Location: Worldwide
- Worldwide238 jobs, applied. Activate to remove
- United States63 jobs
- United Kingdom9 jobs
- Canada5 jobs
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
- English230 jobs
- Germanno roles alongside your other filters
- Spanishno roles alongside your other filters
- French1 job
- Japanese3 jobs
- Portugueseno roles alongside your other filters
Field
- Languages & Linguisticsno roles alongside your other filters
- Audio & Voiceno roles alongside your other filters
- Engineering49 jobs
- Business & Finance30 jobs
- Software & IT32 jobs
- Health & Medicine54 jobs
- Law, Policy & Security35 jobs
- General & Data Collection1 job
- Science & Math41 jobs
- AI Safety & Evaluation20 jobs
- Video, Image & Design3 jobs
- Data, AI & ML13 jobs
- Writing & Education2 jobs
- Other fields5 jobs
Newest
238 open roles matching these filters · page 4 of 10
- $50 – $80 / HourWorldwide
A short, well-paid sprint for very senior software engineers: help a leading foundation-model lab improve its models on hard SWE tasks. 10+ years at top US tech firms, about 20 hours a week for 2–3 weeks. $150–210/hour by geography and level; the 2-hour vetting exercise is paid $100.
$150 – $210 / HourWorldwideRegulatory counsel and compliance lawyers design AI evaluation scenarios, compliance memos, filings and rubrics across financial services, FDA, antitrust, privacy and energy regulation. US (APA, SEC, FDA, FTC) or EU/UK track. Remote hourly contract at $90–100/hour; 5+ years in regulatory practice.
$90 – $100 / HourWorldwidePatent attorneys, patent agents and IP counsel build AI evaluation scenarios on prosecution, licensing, FTO and IP litigation, with reference applications, office action responses and rubrics. US (35 U.S.C., MPEP) or EPC/PCT track. Remote hourly contract at $90–100/hour; 5+ years in IP practice.
$90 – $100 / HourWorldwideDesign GPU programming tasks in CUDA, WebGPU or GLSL for training LLMs on performance and architecture, including kernel profiling and C++ host code. Remote contractor, $60–95/hour paid per accepted task, 25 openings. Graphics, HPC and ML acceleration backgrounds all qualify.
$60 – $95 / HourWorldwideBuild HR evaluation tasks for AI on a US employment-law track, an international (UK/EU) track, or both: scenarios, reference policies and investigation reports, and rubrics. For HR leaders and CHROs with 5+ years at large companies. Remote hourly contract at $70–80/hour.
$70 – $80 / HourWorldwideEngineers with propulsion, initiation or effects test experience red-team frontier AI models: write benign, dual-use and adversarial prompts, judge the replies against a policy standard, and write reference answers with the reasoning. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideUmbrella listing for Mercor's nuclear red-team panel: fuel-cycle engineers, safeguards inspectors, nuclear security, forensics and nonproliferation specialists write prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 21 hired this month.
$65 – $75 / TaskWorldwideBigLaw immigration attorneys review immigration documents and legal scenarios, draft and grade memos and briefs, and correct AI-generated legal output. Ten openings, contractor, fully remote, $140–400/hour. An active licence and complex US immigration casework at a large firm are the stated profile.
$140 – $400 / HourWorldwideBuild AI evaluation tasks set inside Fortune 500 insurance: commercial underwriting, claims adjudication, reserving, reinsurance and NAIC compliance scenarios, with reference guidelines and rubrics. Needs 5+ years at a major carrier or reinsurer. $50–60/hour, remote contract.
$50 – $60 / HourWorldwideMaterials science PhDs author original, executable research problems for a scientific-computing AI benchmark, with depth in both semiconductor materials and molecular modeling. Tasks ship only when frontier models fail them more often than not. 6 weeks, 20+ hours a week, Git and Docker workflow. $70/hour, 212 hired this month.
$70 / HourWorldwideBuild realistic production-management tasks for AI benchmarks (budget reconciliation, schedule changes, vendor and crew coordination) from authentic budgets, call sheets and contracts, then write 35+ criterion rubrics to grade the answers. Five years of credited production experience; paid per accepted task at $45–85/hour. Fifty openings.
$45 – $85 / HourWorldwidePart-time legal AI work for in-house technology lawyers: run simulated negotiations and redlines on MSAs, NDAs and DPAs, grade AI responses and write the evaluation criteria. Requires three years in-house on tech transactions; no bar admission is listed. 50 openings at $85–105/hour.
$85 – $105 / HourWorldwideThe reviewer seat on micro1's physics work: critique derivations and arguments written by researchers or AI, find the errors and unjustified steps, and write precise feedback, cross-checking in SymPy and Python. Postdocs and junior faculty, US, Canada and UK focused. 30 openings, contractor, $80–150/hour.
$80 – $150 / HourWorldwideRadiological emergency planners, field monitoring teams and consequence modellers red-team frontier AI models: write benign, dual-use and adversarial prompts, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideCreate, solve and review hard full-stack engineering tasks, debug across browser, API and database layers, judge AI-generated solutions, and port whole builds from one language to another. Global and fully remote, about 15 flexible hours a week, paid per task on a $50–100/hour band, 300 openings.
$50 – $100 / HourWorldwideA paid expert conversation for engineers who build and run LLM agents in production: a short AI screening interview (no coding), then, if selected, a 30-minute live call on agent reliability, evaluation and internal adoption, paid $100–500 depending on depth of experience.
$100 – $500 / TaskWorldwidePaid pilot for a biotech and pharma research team: create and critique rubrics that assess commercial drugs and development programs, judge investment-style theses, and assess 5–10 companies end to end. For specialist-fund biotech analysts with 5+ years. About 10–20 hours over 1–2 weeks, $120–200/hour.
$120 – $200 / HourWorldwideExplain physics to an AI and grade what it writes back: write clear explanations of hard concepts, build physics training material, and review model output for conceptual accuracy. A physics PhD is the preferred bar. 100 openings, contractor, $70–90/hour, fully remote.
$70 – $90 / HourWorldwideRead US sales and use tax statutes subsection by subsection, write the rubric an AI's formal translation of the law must satisfy, then grade that translation pass/fail and write test scenarios for what it missed. CPA, CA or US tax attorney background required. Ten openings, contract, $20–30/hour.
$20 – $30 / HourWorldwideResearch-level AI training work on magnetic order: magnetic structure factors, propagation vectors, magnetic space groups in BNS notation checked against MAGNDATA, AFM/FM classification, neutron scattering and MOKE signatures. Ten openings, contractor, $80–160/hour, remote.
$80 – $160 / HourWorldwideLend lab chemistry expertise (reactions, synthesis, separations, analytical methods) to AI training data, judging experimental reasoning and explaining it plainly. Contractor, remote, 100 openings, $70–90/hour. A PhD is preferred and hands-on wet-lab experience is expected.
$70 – $90 / HourWorldwideRed-team frontier AI models on chemical safety: write benign, dual-use and adversarial prompts from exposure and process-hazard work, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideLitigate a simulated dispute among shareholders or LLC members turn by turn: draft motions, briefs, demand letters and discovery under senior direction, and write rubric items for each submission. Three years of post-JD litigation on fiduciary duty, oppression, buyout and governance disputes and a US licence (active or lapsed) are the profile. One opening, $100/hour.
$100 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.