Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.

Filter jobs

Location
  • Worldwide34 jobs
  • United States5 jobs
  • United Kingdom1 job
  • Canadano roles alongside your other filters
  • Indiano roles alongside your other filters
  • Mexicono roles alongside your other filters
  • Australiano roles alongside your other filters
  • Belgium1 job
  • Irelandno roles alongside your other filters
  • Argentinano roles alongside your other filters
  • Switzerlandno roles alongside your other filters
  • Chileno roles alongside your other filters
  • Colombiano roles alongside your other filters
  • Costa Ricano roles alongside your other filters
  • Germanyno roles alongside your other filters
  • Spainno roles alongside your other filters
  • Panamano roles alongside your other filters
  • Brazilno roles alongside your other filters
  • Dominican Republicno roles alongside your other filters
  • Franceno roles alongside your other filters
  • Italyno roles alongside your other filters
  • Luxembourgno roles alongside your other filters
  • Netherlandsno roles alongside your other filters
  • New Zealandno roles alongside your other filters
  • Peruno roles alongside your other filters
  • Uruguayno roles alongside your other filters
  • Albaniano roles alongside your other filters
  • Austriano roles alongside your other filters
  • Bosnia & Herzegovinano roles alongside your other filters
  • Barbadosno roles alongside your other filters
  • Bulgariano roles alongside your other filters
  • Bahamasno roles alongside your other filters
  • Czechiano roles alongside your other filters
  • Denmarkno roles alongside your other filters
  • Ecuadorno roles alongside your other filters
  • Estoniano roles alongside your other filters
  • Finlandno roles alongside your other filters
  • Greeceno roles alongside your other filters
  • Guatemalano roles alongside your other filters
  • Hondurasno roles alongside your other filters
  • Croatiano roles alongside your other filters
  • Hungaryno roles alongside your other filters
  • Icelandno roles alongside your other filters
  • Jamaicano roles alongside your other filters
  • South Koreano roles alongside your other filters
  • Liechtensteinno roles alongside your other filters
  • Lithuaniano roles alongside your other filters
  • Latviano roles alongside your other filters
  • Monacono roles alongside your other filters
  • Moldovano roles alongside your other filters
  • North Macedoniano roles alongside your other filters
  • Maltano roles alongside your other filters
  • Nicaraguano roles alongside your other filters
  • Norwayno roles alongside your other filters
  • Polandno roles alongside your other filters
  • Portugalno roles alongside your other filters
  • Romaniano roles alongside your other filters
  • Serbiano roles alongside your other filters
  • Swedenno roles alongside your other filters
  • Sloveniano roles alongside your other filters
  • Slovakiano roles alongside your other filters
  • San Marinono roles alongside your other filters
  • El Salvadorno roles alongside your other filters
  • Kosovono roles alongside your other filters
  • Boliviano roles alongside your other filters
  • Belizeno roles alongside your other filters
  • Cubano roles alongside your other filters
  • Indonesiano roles alongside your other filters
  • Japanno roles alongside your other filters
  • Paraguayno roles alongside your other filters
  • Venezuelano roles alongside your other filters
  • South Africano roles alongside your other filters
Show fewer
Language
Show all 49 language options
Field: Science & Math
Level

41 open roles matching these filters · page 1 of 2

  • Write and review Lean 4 proofs for a leading AI lab, formalize informal mathematics and judge whether a model's proof actually proves the right statement. A W-2 part-time employment position through Cincinnatus LLC, at least 20 hours a week and up to 40, paying $90–110/hour.

    $90 – $110 / HourWorldwide
  • Condensed matter PhDs create, solve, review or audit research-level problems for CritPt, a public benchmark testing whether frontier AI models can do real physics research. Nineteen narrow research areas, from bosonization to SYK. Remote hourly contract at $80–110 per hour.

    $80 – $110 / HourWorldwide
  • AMO physicists create, solve, review or audit research-level problems for CritPt, a public benchmark testing whether frontier AI models can do real physics research. Seven narrow areas, including levitated optomechanics, cavity QED and ultracold atoms in optical lattices. Remote hourly contract at $80–110 per hour.

    $80 – $110 / HourWorldwide
  • Researchers who have published on stochastic autocatalytic growth create, solve, review or audit research-level problems for CritPt, a public AI physics benchmark. A single narrow area: chemical master equations, branching processes and reaction-network moments. Remote hourly contract at $80–110 per hour.

    $80 – $110 / HourWorldwide
  • High energy and nuclear theorists create, solve, review or audit research-level problems for CritPt, a public benchmark testing whether frontier AI models can do real physics research. Seven narrow areas, from AdS/BCFT to quasi-PDFs and dark photon searches. Remote hourly contract at $80–110 per hour.

    $80 – $110 / HourWorldwide
  • Mathematical physicists create, solve, review or audit research-level problems for CritPt, a public AI physics benchmark, where the standard of proof sits closer to mathematics than physics. Four narrow areas, from hypergeometric identities to Fefferman-Graham geometry. Remote hourly contract at $80–110 per hour.

    $80 – $110 / HourWorldwide
  • Remote hourly contract for experimental physicists, chemists and materials scientists who already use Origin or OriginPro on their own Windows PC for graphing and curve fitting. $45–55/hour, paid weekly via Stripe or Wise. Bring your own licence and a screen above 2.5 megapixels.

    $45 – $55 / HourWorldwide
  • Build and evaluate training data for a frontier lab's materials science models: DFT, AIMD, classical MD, surface and adsorption modeling, reaction energetics. For US-based computational PhDs fluent in VASP, Quantum ESPRESSO, CP2K, LAMMPS or ASE. Long-term, 10–40 hours a week, $84/hour.

    $84 / HourOpen to United States
  • Materials science PhDs author original, executable research problems for a scientific-computing AI benchmark, with depth in both semiconductor materials and molecular modeling. Tasks ship only when frontier models fail them more often than not. 6 weeks, 20+ hours a week, Git and Docker workflow. $70/hour, 212 hired this month.

    $70 / HourWorldwide
  • Ukrainian-speaking PhD chemists and biologists write specialised science prompts in Ukrainian and grade AI answers for accuracy and safe handling of dual-use topics. Part-time remote at $48–52/hour, $10 above the Ukrainian generalist role. Eastern Europe preferred, not required; 11 hires this month.

    $48 – $52 / HourWorldwide
  • Japanese-speaking PhD chemists and biologists write specialised science prompts in Japanese and grade AI answers for accuracy and dual-use safety. Part-time remote at $68–72/hour, second-highest in the series and $20 above the Japanese generalist role. East Asia preferred, not required.

    $68 – $72 / HourWorldwide
  • Red-team frontier AI models on chemical safety: write benign, dual-use and adversarial prompts from exposure and process-hazard work, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Full-time Bay Area hybrid role with an AI lab's research team: review model reasoning on life sciences tasks, write golden solutions and instruction specs, and design benchmarks. For life sciences PhDs with 4+ years of substantive research experience. W-2 via Cincinnatus, $65–105/hour.

    $65 – $105 / HourHybridOpen to United States
  • Umbrella listing for Mercor's chemistry red-team panel: synthetic, analytical, forensic, defence and process safety chemists write benign, dual-use and adversarial prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 5 hired this month.

    $65 – $75 / TaskWorldwide
  • Red-team frontier AI models on chemical misuse: write benign, dual-use and adversarial prompts from trace analysis and method validation, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Red-team frontier AI models on chemical misuse: write benign, dual-use and adversarial prompts from route design and scale-up experience, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Write and verify expert-level physics multiple-choice questions (one right answer, nine plausible wrong ones) for an AI benchmark, across semiconductors, photonics, quantum sensing, plasma, turbulence and geophysics. PhD or doctoral candidate preferred. Remote hourly contract at $61–77/hour, 10+ hours a week.

    $61 – $77 / HourWorldwide
  • PhD chemists and biologists who write fluent Finnish: author specialised science prompts and judge how AI models answer them, including where a question strays into dual-use territory. Part-time remote at $61–65/hour, $13 above the Finnish generalist role. Finland or Western Europe preferred, not required.

    $61 – $65 / HourWorldwide
  • Thai-speaking PhD chemists and biologists write specialised science prompts in Thai and grade how AI models answer them, with a focus on dual-use safety. Part-time remote at $24–28/hour, $6 above the Thai generalist role. Southeast Asia preferred, not required; 10 hires this month.

    $24 – $28 / HourWorldwide
  • For Danish-speaking PhD scientists in chemistry or biology: write specialised prompts in Danish, grade AI answers for accuracy and safe handling, and classify conversations against guidelines. Part-time remote at $61–65/hour. Denmark or Western Europe preferred, not required; PhD candidates eligible.

    $61 – $65 / HourWorldwide
  • Formulation and synthesis chemists from pyrotechnics or propellant work red-team frontier AI models: write benign, dual-use and adversarial prompts, grade the model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • The best-paid role in Mercor's bilingual AI safety series: Norwegian-speaking PhD chemists and biologists write specialised science prompts and grade how AI models handle dual-use questions. Part-time remote at $77–81/hour. Norway preferred, not required; PhD candidates eligible.

    $77 – $81 / HourWorldwide
  • Experimental scientists create and review training data for a frontier lab's materials science models: inorganic synthesis, superconductors, semiconductors and advanced packaging, characterization (XRD, SEM, TEM) and fabrication. PhD, MS or equivalent hands-on experience. US-based, 10–40 hours a week, $84/hour.

    $84 / HourOpen to United States
  • Full-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior drug development scientists write golden solutions, instruction specs and benchmarks for pharma R&D reasoning. Hybrid in the Bay Area, 40 hours a week for an initial 6 months. PhD, PharmD or MD with 4+ years in industry R&D. $75–115/hour.

    $75 – $115 / HourHybridOpen to United States

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.