Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • Research-level benchmark work on bacterial population growth: two-state growth-rate switching with gamma-distributed waiting times, the Euler-Lotka equation, renewal theory and first-passage times. Solver, Auditor or Adjudicator roles. Ten openings, contractor, $80–160/hour, remote.

    $80 – $160 / HourWorldwide
  • Seasoned transactional lawyers redline contracts in simulated negotiations, grade AI responses to contract scenarios, and design the criteria that score them. Three years in-house on technology transactions (MSAs, NDAs, ISDAs, derivatives) is the stated bar; M&A or fund formation at a firm is preferred. 100 openings, task-based at $90–130/hour and about 3.5 hours per task.

    $90 – $130 / HourWorldwide
  • Evaluate AI-generated psychological content for clinical accuracy, bias and cultural fit, and build case studies and therapeutic dialogues for training. PhD in psychology or MD with psychiatry residency, plus native fluency in one of 16 listed languages and advanced English. 70 openings, contractor, $100–200/hour.

    $100 – $200 / HourWorldwide
  • Curate and annotate drug-discovery datasets, judge AI answers on medicinal chemistry and molecular analysis, and write feedback that improves the model. Contractor, remote, 30 openings, $90–120/hour. An advanced degree and cheminformatics or omics experience are the preferred background.

    $90 – $120 / HourWorldwide
  • Act as ground truth for an AI lab teaching models real enterprise sales work: audit workflows, build golden reference trajectories in a mock sales stack, and refine rubrics. For sellers with around 10 years in enterprise sales. US only, $60–90/hour, placed via Cincinnatus.

    $60 – $90 / HourOpen to United States
  • Senior mechanical engineers from Fortune 500 or major industrial manufacturers write AI evaluation scenarios drawn from real product design, analysis and certification work, with reference outputs and rubrics on an ASME or ISO track. 5+ years required. Remote hourly contract at $70–80/hour; 285 hires this month.

    $70 – $80 / HourWorldwide
  • Formulation and synthesis chemists from pyrotechnics or propellant work red-team frontier AI models: write benign, dual-use and adversarial prompts, grade the model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Senior commercial litigators build AI evaluation scenarios covering discovery, motions, trial and arbitration, with reference pleadings, motions and rubrics. US (FRCP, FRE) or international (English CPR, ICC, LCIA) track. Remote hourly contract at $90–100/hour; 5+ years with trial or arbitration experience.

    $90 – $100 / HourWorldwide
  • Certified bomb technicians, EOD veterans and bomb squad leaders red-team frontier AI models: write benign, dual-use and adversarial prompts from public-safety practice, judge model responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Redline contracts in simulated negotiations, grade how an AI handles the same scenarios, and write the criteria that score it. This In-House Counsel variant asks for three years in-house on technology transactions including ISDAs and derivatives. 100 openings, task-based pay at $90–130/hour and roughly 3.5 hours per task.

    $90 – $130 / HourWorldwide
  • Experimental scientists create and review training data for a frontier lab's materials science models: inorganic synthesis, superconductors, semiconductors and advanced packaging, characterization (XRD, SEM, TEM) and fabrication. PhD, MS or equivalent hands-on experience. US-based, 10–40 hours a week, $84/hour.

    $84 / HourOpen to United States
  • Hands-on ML researchers take on scoped, open-ended empirical problems: training image classifiers and generators from scratch, fine-tuning open-weight LLMs, adversarial robustness, compression under hard budgets, and multilingual pre-training. 3+ years of ML research (PhD counts). Remote hourly contract at $100–120/hour.

    $100 – $120 / HourWorldwide
  • Export control, treaty and proliferation analysts red-team frontier AI models: write benign, dual-use and adversarial prompts from nonproliferation work, judge model responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Full-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior drug development scientists write golden solutions, instruction specs and benchmarks for pharma R&D reasoning. Hybrid in the Bay Area, 40 hours a week for an initial 6 months. PhD, PharmD or MD with 4+ years in industry R&D. $75–115/hour.

    $75 – $115 / HourHybridOpen to United States
  • A broad legal AI training brief for lawyers with Vault Law 100 experience: analyse US law questions, research, draft and edit memos and contracts, and quality-check legal content. Firm pedigree is mandatory; practice area is open. 100 openings at $140–400/hour, remote contractor.

    $140 – $400 / HourWorldwide
  • Nuclear forensics, radiochemistry and detection specialists red-team frontier AI models: write benign, dual-use and adversarial prompts from characterisation and attribution work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • A full-time W-2 role (via Cincinnatus LLC) embedded with a leading AI lab in the Bay Area: review legal model outputs, write instruction specs and golden solutions, and build legal benchmarks. Hybrid, on-site several days a week, 6-month initial term. $60–100/hour; JD, 5+ years' practice, US bar.

    $60 – $100 / HourHybridOpen to United States
  • MC&A, physical protection and vulnerability assessment specialists red-team frontier AI models: write benign, dual-use and adversarial prompts from nuclear security work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Author or verify expert multiple-choice biology questions (one correct answer, nine subtle distractors, chain-of-thought solution, references) for an AI benchmark in pharma manufacturing, synthetic biology, drug discovery and agricultural, environmental and food biology. PhD or candidate preferred. $60–75/hour, 10+ hours a week.

    $60 – $75 / HourWorldwide
  • Author point-in-time election and political-risk forecasts, document calls on polling and win probabilities, and grade AI analyses. For senior national forecasters, campaign analytics leads, pollsters and political-risk analysts with 8+ years. Remote, $150–250/hour.

    $150 – $250 / HourWorldwide
  • Map how an upcoming catalyst should ripple through suppliers, customers, competitors and substitutes, estimate direction and magnitude from primary filings, and grade AI analyses. For senior sector PMs and lead equity analysts with 8+ years. Remote, $150–250/hour.

    $150 – $250 / HourWorldwide
  • Write a Python simulation and a spec sheet with pass/fail thresholds; a frontier model probes your simulation a limited number of times, then submits a design that an agentic grader scores. For control, analog or RF circuit, power electronics or mechanical design experts with a PhD or equivalent industry record. Remote hourly contract, $60–90/hour.

    $60 – $90 / HourWorldwide
  • Quantum and computational chemistry PhDs write original, runnable research problems for a scientific-computing AI benchmark, with grading criteria, calibrated until frontier models fail them more often than they pass. 6 weeks at 20+ hours a week, Git and Docker workflow. Flat $70/hour; 182 hired this month.

    $70 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.