Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.

Filter jobs

Location
  • Worldwide23 jobs
  • United States6 jobs
  • United Kingdomno roles alongside your other filters
  • Canadano roles alongside your other filters
  • Indiano roles alongside your other filters
  • Mexicono roles alongside your other filters
Show all 72 location options
Language
  • English29 jobs
  • Germanno roles alongside your other filters
  • Spanishno roles alongside your other filters
  • Frenchno roles alongside your other filters
  • Japaneseno roles alongside your other filters
  • Portugueseno roles alongside your other filters
Show all 49 language options
Field: Engineering
Level: Senior

29 open roles matching these filters · page 1 of 2

  • Hands-on structural, thermal, mechanical design and dynamics engineers review and write hard engineering problems about real hardware (loads, margins, heat transfer, vibration, tolerances) for an AI research initiative. 5+ years with 3 recent hands-on. Remote contract, $100–120/hour.

    $100 – $120 / HourWorldwide
  • Hands-on systems, integration, reliability and manufacturing test engineers review and write hard engineering problems about real hardware (V&V, qualification, FMEA, root-cause analysis) for an AI research initiative. 5+ years with 3 recent hands-on. Remote contract, $100–120/hour, weekly via Stripe or Wise.

    $100 – $120 / HourWorldwide
  • Embedded firmware, FPGA/RTL, flight software and hardware test automation engineers review and write hard problems about software that runs on real hardware (timing, race conditions, interfaces, bring-up) for an AI research initiative. 5+ years, 3 recent hands-on. Remote contract, $100–120/hour.

    $100 – $120 / HourWorldwide
  • Hands-on RF, power electronics, analog/mixed-signal, PCB and signal integrity engineers review and write hard problems about real hardware designs, measurements and bring-up for an AI research initiative. 5+ years with 3 recent hands-on. Remote contract, $100–120/hour, weekly via Stripe or Wise.

    $100 – $120 / HourWorldwide
  • Senior civil engineers build infrastructure evaluation tasks for AI: realistic design, permitting and construction scenarios, reference calculations and rubrics, on either a US (ASCE, ACI, AISC, AASHTO) or International (Eurocodes, ISO) standards track. 5+ years, PE or equivalent strongly preferred. Remote hourly contract at $70–80/hour.

    $70 – $80 / HourWorldwide
  • Generate, structure and evaluate expert ALD and thin-film data for a frontier AI lab building semiconductor and physical-science models: solve hard problems, rate model reasoning, and turn recipes into model-ready data. US-based, 10–40 hours a week, $84/hour.

    $84 / HourOpen to United States
  • Umbrella listing for Mercor's nuclear red-team panel: fuel-cycle engineers, safeguards inspectors, nuclear security, forensics and nonproliferation specialists write prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 21 hired this month.

    $65 – $75 / TaskWorldwide
  • Build and evaluate training data for a frontier lab's materials science models: DFT, AIMD, classical MD, surface and adsorption modeling, reaction energetics. For US-based computational PhDs fluent in VASP, Quantum ESPRESSO, CP2K, LAMMPS or ASE. Long-term, 10–40 hours a week, $84/hour.

    $84 / HourOpen to United States
  • Materials science PhDs author original, executable research problems for a scientific-computing AI benchmark, with depth in both semiconductor materials and molecular modeling. Tasks ship only when frontier models fail them more often than not. 6 weeks, 20+ hours a week, Git and Docker workflow. $70/hour, 212 hired this month.

    $70 / HourWorldwide
  • Radiological emergency planners, field monitoring teams and consequence modellers red-team frontier AI models: write benign, dual-use and adversarial prompts, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Author AI evaluation tasks from real fire and life safety review work: egress plan checks, sprinkler and alarm review, hazmat control areas, firestop photo verification. For US fire marshals, fire protection engineers and NICET III+ designer-reviewers with 3+ years in the seat. Remote hourly contract at $45–60/hour.

    $45 – $60 / HourOpen to United States
  • Senior electrical engineers from large technology, industrial or energy companies build AI evaluation scenarios across power systems, PCB and semiconductor design, embedded firmware, signal processing and controls, with reference analyses and rubrics on a US (NEC, NESC, IEEE) or IEC track. Remote hourly contract at $70–80/hour.

    $70 – $80 / HourWorldwide
  • A full-time W-2 placement at a leading AI lab through Cincinnatus LLC: US-based mechanical engineers vet model outputs, write instruction specs and reference solutions, and build benchmarks. 5+ years in industry and a mechanical engineering degree required. 40 hours a week for an initial 2–3 months, $60–90/hour.

    $60 – $90 / HourOpen to United States
  • Senior mechanical engineers from Fortune 500 or major industrial manufacturers write AI evaluation scenarios drawn from real product design, analysis and certification work, with reference outputs and rubrics on an ASME or ISO track. 5+ years required. Remote hourly contract at $70–80/hour; 285 hires this month.

    $70 – $80 / HourWorldwide
  • Experimental scientists create and review training data for a frontier lab's materials science models: inorganic synthesis, superconductors, semiconductors and advanced packaging, characterization (XRD, SEM, TEM) and fabrication. PhD, MS or equivalent hands-on experience. US-based, 10–40 hours a week, $84/hour.

    $84 / HourOpen to United States
  • Export control, treaty and proliferation analysts red-team frontier AI models: write benign, dual-use and adversarial prompts from nonproliferation work, judge model responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Nuclear forensics, radiochemistry and detection specialists red-team frontier AI models: write benign, dual-use and adversarial prompts from characterisation and attribution work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • MC&A, physical protection and vulnerability assessment specialists red-team frontier AI models: write benign, dual-use and adversarial prompts from nuclear security work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Write a Python simulation and a spec sheet with pass/fail thresholds; a frontier model probes your simulation a limited number of times, then submits a design that an agentic grader scores. For control, analog or RF circuit, power electronics or mechanical design experts with a PhD or equivalent industry record. Remote hourly contract, $60–90/hour.

    $60 – $90 / HourWorldwide
  • Fuel-cycle and criticality safety engineers red-team frontier AI models: write benign, dual-use and adversarial prompts from enrichment, fabrication, reprocessing or criticality work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Current and former IAEA, Euratom and state-system safeguards inspectors and accountancy analysts red-team frontier AI models: write benign, dual-use and adversarial prompts, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Nuclear medicine physicists, radiopharmacy staff and cyclotron RSOs red-team frontier AI models: write benign, dual-use and adversarial prompts from medical isotope work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Radioactive source security and vulnerability assessment specialists red-team frontier AI models: write benign, dual-use and adversarial prompts from Category 1 and 2 source security work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Full-time Bay Area hybrid role embedded with an AI lab: review model reasoning on materials problems, write golden solutions and specs, and build benchmarks. For materials PhDs (or master's with exceptional industrial depth) with 4+ years of R&D. W-2 via Cincinnatus, $70–110/hour.

    $70 – $110 / HourHybridOpen to United States

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.