Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • Write and verify expert-level physics multiple-choice questions (one right answer, nine plausible wrong ones) for an AI benchmark, across semiconductors, photonics, quantum sensing, plasma, turbulence and geophysics. PhD or doctoral candidate preferred. Remote hourly contract at $61–77/hour, 10+ hours a week.

    $61 – $77 / HourWorldwide
  • Design Fortune 500 go-to-market scenarios, write reference deal strategies and campaigns, and author rubrics that test whether AI shows real enterprise sales and marketing judgment. For senior F500 sellers, marketers and RevOps leads. Remote hourly contract at $60–70/hour.

    $60 – $70 / HourWorldwide
  • Licensed blasters and blasting engineers red-team frontier AI models: write benign, dual-use and adversarial prompts from quarry, mine and demolition practice, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Senior electrical engineers from large technology, industrial or energy companies build AI evaluation scenarios across power systems, PCB and semiconductor design, embedded firmware, signal processing and controls, with reference analyses and rubrics on a US (NEC, NESC, IEEE) or IEC track. Remote hourly contract at $70–80/hour.

    $70 – $80 / HourWorldwide
  • Civil rights, environmental and legal aid attorneys build AI evaluation scenarios, reference briefs and rubrics on a US track (civil rights statutes, NEPA, CAA, CWA) or an international one (human rights law, EU directives). Remote hourly contract at $90–100/hour; 5+ years in public interest practice.

    $90 – $100 / HourWorldwide
  • A full-time W-2 placement at a leading AI lab through Cincinnatus LLC: US-based mechanical engineers vet model outputs, write instruction specs and reference solutions, and build benchmarks. 5+ years in industry and a mechanical engineering degree required. 40 hours a week for an initial 2–3 months, $60–90/hour.

    $60 – $90 / HourOpen to United States
  • Judge AI-enhanced and upscaled video and stills at pixel level for a leading AI lab's GenAI team: artifacts, noise, aliasing, banding, grain, sharpening. For VFX and rendering supervisors, colorists, DPs, lighting artists and high-end photographers with 5+ years. US, 20 hours a week, $60–90/hour.

    $60 – $90 / HourOpen to United States
  • Act as ground truth for an AI lab teaching models real enterprise sales work: audit workflows, build golden reference trajectories in a mock sales stack, and refine rubrics. For sellers with around 10 years in enterprise sales. US only, $60–90/hour, placed via Cincinnatus.

    $60 – $90 / HourOpen to United States
  • Senior mechanical engineers from Fortune 500 or major industrial manufacturers write AI evaluation scenarios drawn from real product design, analysis and certification work, with reference outputs and rubrics on an ASME or ISO track. 5+ years required. Remote hourly contract at $70–80/hour; 285 hires this month.

    $70 – $80 / HourWorldwide
  • Formulation and synthesis chemists from pyrotechnics or propellant work red-team frontier AI models: write benign, dual-use and adversarial prompts, grade the model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Senior commercial litigators build AI evaluation scenarios covering discovery, motions, trial and arbitration, with reference pleadings, motions and rubrics. US (FRCP, FRE) or international (English CPR, ICC, LCIA) track. Remote hourly contract at $90–100/hour; 5+ years with trial or arbitration experience.

    $90 – $100 / HourWorldwide
  • Certified bomb technicians, EOD veterans and bomb squad leaders red-team frontier AI models: write benign, dual-use and adversarial prompts from public-safety practice, judge model responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Experimental scientists create and review training data for a frontier lab's materials science models: inorganic synthesis, superconductors, semiconductors and advanced packaging, characterization (XRD, SEM, TEM) and fabrication. PhD, MS or equivalent hands-on experience. US-based, 10–40 hours a week, $84/hour.

    $84 / HourOpen to United States
  • Hands-on ML researchers take on scoped, open-ended empirical problems: training image classifiers and generators from scratch, fine-tuning open-weight LLMs, adversarial robustness, compression under hard budgets, and multilingual pre-training. 3+ years of ML research (PhD counts). Remote hourly contract at $100–120/hour.

    $100 – $120 / HourWorldwide
  • Export control, treaty and proliferation analysts red-team frontier AI models: write benign, dual-use and adversarial prompts from nonproliferation work, judge model responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Full-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior drug development scientists write golden solutions, instruction specs and benchmarks for pharma R&D reasoning. Hybrid in the Bay Area, 40 hours a week for an initial 6 months. PhD, PharmD or MD with 4+ years in industry R&D. $75–115/hour.

    $75 – $115 / HourHybridOpen to United States
  • Nuclear forensics, radiochemistry and detection specialists red-team frontier AI models: write benign, dual-use and adversarial prompts from characterisation and attribution work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • A full-time W-2 role (via Cincinnatus LLC) embedded with a leading AI lab in the Bay Area: review legal model outputs, write instruction specs and golden solutions, and build legal benchmarks. Hybrid, on-site several days a week, 6-month initial term. $60–100/hour; JD, 5+ years' practice, US bar.

    $60 – $100 / HourHybridOpen to United States
  • MC&A, physical protection and vulnerability assessment specialists red-team frontier AI models: write benign, dual-use and adversarial prompts from nuclear security work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Author or verify expert multiple-choice biology questions (one correct answer, nine subtle distractors, chain-of-thought solution, references) for an AI benchmark in pharma manufacturing, synthetic biology, drug discovery and agricultural, environmental and food biology. PhD or candidate preferred. $60–75/hour, 10+ hours a week.

    $60 – $75 / HourWorldwide
  • Author point-in-time election and political-risk forecasts, document calls on polling and win probabilities, and grade AI analyses. For senior national forecasters, campaign analytics leads, pollsters and political-risk analysts with 8+ years. Remote, $150–250/hour.

    $150 – $250 / HourWorldwide
  • Map how an upcoming catalyst should ripple through suppliers, customers, competitors and substitutes, estimate direction and magnitude from primary filings, and grade AI analyses. For senior sector PMs and lead equity analysts with 8+ years. Remote, $150–250/hour.

    $150 – $250 / HourWorldwide
  • Write a Python simulation and a spec sheet with pass/fail thresholds; a frontier model probes your simulation a limited number of times, then submits a design that an agentic grader scores. For control, analog or RF circuit, power electronics or mechanical design experts with a PhD or equivalent industry record. Remote hourly contract, $60–90/hour.

    $60 – $90 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.