Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.

238 open roles matching these filters · page 4 of 10

  • Write and review evaluation tasks built on journal manuscripts, abstracts, posters and decks, and judge AI-drafted sections against ICMJE, GPP and EQUATOR guidelines and the source CSRs and TFLs. Five years of publication authoring at a sponsor, CRO or MedComms agency. Fifty openings, contractor, $50–80/hour.

    $50 – $80 / HourWorldwide
  • A short, well-paid sprint for very senior software engineers: help a leading foundation-model lab improve its models on hard SWE tasks. 10+ years at top US tech firms, about 20 hours a week for 2–3 weeks. $150–210/hour by geography and level; the 2-hour vetting exercise is paid $100.

    $150 – $210 / HourWorldwide
  • Regulatory counsel and compliance lawyers design AI evaluation scenarios, compliance memos, filings and rubrics across financial services, FDA, antitrust, privacy and energy regulation. US (APA, SEC, FDA, FTC) or EU/UK track. Remote hourly contract at $90–100/hour; 5+ years in regulatory practice.

    $90 – $100 / HourWorldwide
  • Patent attorneys, patent agents and IP counsel build AI evaluation scenarios on prosecution, licensing, FTO and IP litigation, with reference applications, office action responses and rubrics. US (35 U.S.C., MPEP) or EPC/PCT track. Remote hourly contract at $90–100/hour; 5+ years in IP practice.

    $90 – $100 / HourWorldwide
  • Design GPU programming tasks in CUDA, WebGPU or GLSL for training LLMs on performance and architecture, including kernel profiling and C++ host code. Remote contractor, $60–95/hour paid per accepted task, 25 openings. Graphics, HPC and ML acceleration backgrounds all qualify.

    $60 – $95 / HourWorldwide
  • Build HR evaluation tasks for AI on a US employment-law track, an international (UK/EU) track, or both: scenarios, reference policies and investigation reports, and rubrics. For HR leaders and CHROs with 5+ years at large companies. Remote hourly contract at $70–80/hour.

    $70 – $80 / HourWorldwide
  • Engineers with propulsion, initiation or effects test experience red-team frontier AI models: write benign, dual-use and adversarial prompts, judge the replies against a policy standard, and write reference answers with the reasoning. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Umbrella listing for Mercor's nuclear red-team panel: fuel-cycle engineers, safeguards inspectors, nuclear security, forensics and nonproliferation specialists write prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 21 hired this month.

    $65 – $75 / TaskWorldwide
  • BigLaw immigration attorneys review immigration documents and legal scenarios, draft and grade memos and briefs, and correct AI-generated legal output. Ten openings, contractor, fully remote, $140–400/hour. An active licence and complex US immigration casework at a large firm are the stated profile.

    $140 – $400 / HourWorldwide
  • Build AI evaluation tasks set inside Fortune 500 insurance: commercial underwriting, claims adjudication, reserving, reinsurance and NAIC compliance scenarios, with reference guidelines and rubrics. Needs 5+ years at a major carrier or reinsurer. $50–60/hour, remote contract.

    $50 – $60 / HourWorldwide
  • Materials science PhDs author original, executable research problems for a scientific-computing AI benchmark, with depth in both semiconductor materials and molecular modeling. Tasks ship only when frontier models fail them more often than not. 6 weeks, 20+ hours a week, Git and Docker workflow. $70/hour, 212 hired this month.

    $70 / HourWorldwide
  • Build realistic production-management tasks for AI benchmarks (budget reconciliation, schedule changes, vendor and crew coordination) from authentic budgets, call sheets and contracts, then write 35+ criterion rubrics to grade the answers. Five years of credited production experience; paid per accepted task at $45–85/hour. Fifty openings.

    $45 – $85 / HourWorldwide
  • Part-time legal AI work for in-house technology lawyers: run simulated negotiations and redlines on MSAs, NDAs and DPAs, grade AI responses and write the evaluation criteria. Requires three years in-house on tech transactions; no bar admission is listed. 50 openings at $85–105/hour.

    $85 – $105 / HourWorldwide
  • The reviewer seat on micro1's physics work: critique derivations and arguments written by researchers or AI, find the errors and unjustified steps, and write precise feedback, cross-checking in SymPy and Python. Postdocs and junior faculty, US, Canada and UK focused. 30 openings, contractor, $80–150/hour.

    $80 – $150 / HourWorldwide
  • Radiological emergency planners, field monitoring teams and consequence modellers red-team frontier AI models: write benign, dual-use and adversarial prompts, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Create, solve and review hard full-stack engineering tasks, debug across browser, API and database layers, judge AI-generated solutions, and port whole builds from one language to another. Global and fully remote, about 15 flexible hours a week, paid per task on a $50–100/hour band, 300 openings.

    $50 – $100 / HourWorldwide
  • A paid expert conversation for engineers who build and run LLM agents in production: a short AI screening interview (no coding), then, if selected, a 30-minute live call on agent reliability, evaluation and internal adoption, paid $100–500 depending on depth of experience.

    $100 – $500 / TaskWorldwide
  • Paid pilot for a biotech and pharma research team: create and critique rubrics that assess commercial drugs and development programs, judge investment-style theses, and assess 5–10 companies end to end. For specialist-fund biotech analysts with 5+ years. About 10–20 hours over 1–2 weeks, $120–200/hour.

    $120 – $200 / HourWorldwide
  • Explain physics to an AI and grade what it writes back: write clear explanations of hard concepts, build physics training material, and review model output for conceptual accuracy. A physics PhD is the preferred bar. 100 openings, contractor, $70–90/hour, fully remote.

    $70 – $90 / HourWorldwide
  • Read US sales and use tax statutes subsection by subsection, write the rubric an AI's formal translation of the law must satisfy, then grade that translation pass/fail and write test scenarios for what it missed. CPA, CA or US tax attorney background required. Ten openings, contract, $20–30/hour.

    $20 – $30 / HourWorldwide
  • Research-level AI training work on magnetic order: magnetic structure factors, propagation vectors, magnetic space groups in BNS notation checked against MAGNDATA, AFM/FM classification, neutron scattering and MOKE signatures. Ten openings, contractor, $80–160/hour, remote.

    $80 – $160 / HourWorldwide
  • Lend lab chemistry expertise (reactions, synthesis, separations, analytical methods) to AI training data, judging experimental reasoning and explaining it plainly. Contractor, remote, 100 openings, $70–90/hour. A PhD is preferred and hands-on wet-lab experience is expected.

    $70 – $90 / HourWorldwide
  • Red-team frontier AI models on chemical safety: write benign, dual-use and adversarial prompts from exposure and process-hazard work, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Litigate a simulated dispute among shareholders or LLC members turn by turn: draft motions, briefs, demand letters and discovery under senior direction, and write rubric items for each submission. Three years of post-JD litigation on fiduciary duty, oppression, buyout and governance disputes and a US licence (active or lapsed) are the profile. One opening, $100/hour.

    $100 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.