Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • Experienced civil engineers write realistic evaluation tasks for AI (structural, transportation, water, site development), build the supporting documents and data files, solve each task and write multi-item grading rubrics. Paid per accepted task against a $50–100/hour band. Remote, 50 openings.

    $50 – $100 / HourWorldwide
  • Civil and structural engineers author full drawing sets and BIM models in Revit, ArchiCAD or BlenderBIM, export and validate IFC files, and document design decisions for AI training. Five years' practice and three years of daily BIM required. Remote contractor, 10 openings, $30–50/hour.

    $30 – $50 / HourWorldwide
  • Write and review evaluation tasks from CSRs, safety reports and regulatory correspondence in oncology and haematology, judging efficacy and safety conclusions, dose escalation, AE grading and benefit-risk reasoning. Board-certified oncologists or haematologists with five years post-training and trial experience. Thirty openings, $160–200/hour.

    $160 – $200 / HourWorldwide
  • GPU specialists profile and optimise CUDA kernels, refactor C++ and CUDA code, and write GLSL and WebGPU shaders on a project run with a leading AI lab. Remote contractor role, 50 openings, $60–100/hour.

    $60 – $100 / HourWorldwide
  • Write and review evaluation tasks from MedEd modules, advisory board summaries, slide decks and HCP materials, and judge AI-generated MedComms for accuracy, fair balance, audience framing and source fidelity. Five years in medical communications, MedEd or medical affairs. Fifty openings, contractor, $50–80/hour.

    $50 – $80 / HourWorldwide
  • Build securities and commodities scenarios for AI agents from real market data, client files and regulation, decide the compliant answer, and write 35+ point rubrics. Remote contractor, $30–80/hour paid per accepted task, 50 openings. Series 7/63 or equivalent and 4+ years preferred.

    $30 – $80 / HourWorldwide
  • Review inpatient cases for clinical quality, documentation accuracy and appropriate care, make the call on unclear or unusual cases, and help shape the review standards used to train AI. MD or DO with ten years of attending hospitalist experience. Thirty openings, contractor, $70–100/hour.

    $70 – $100 / HourWorldwide
  • Build reinforcement learning environments that test whether AI models can fix bugs, add features, refactor and optimise real code, each with a reproducible setup and a golden reference solution. Around 15 hours a week, paid per accepted task on a $100–150/hour band, 100 openings, no AI experience needed.

    $100 – $150 / HourWorldwide
  • Package real bug fixes, features, refactors and performance problems as reproducible reinforcement learning environments with golden reference solutions for AI coding models. About 15 hours a week, paid per accepted task on a $50–100/hour band, 100 openings. Same text as a higher-paying sibling posting.

    $50 – $100 / HourWorldwide
  • Write and review evaluation tasks that make an AI derive, reproduce or validate clinical trial statistics from TFLs and CDISC datasets, and catch mismatches between outputs, the SAP and the CSR narrative. Remote contractor, $60–65/hour, 30 openings; 5+ years in clinical trials and SAS or R preferred.

    $60 – $65 / HourWorldwide
  • Solve advanced mechanical engineering problems and judge AI-generated answers across design, structural and thermal analysis, FEA, vibration and R&D. Remote contractor paid per accepted task on a $60–100/hour band, 25 openings. MS/PhD, or BS plus 8+ years; problem-authoring credentials strongly preferred.

    $60 – $100 / HourWorldwide
  • Author AI evaluation tasks from real fire and life safety review work: egress plan checks, sprinkler and alarm review, hazmat control areas, firestop photo verification. For US fire marshals, fire protection engineers and NICET III+ designer-reviewers with 3+ years in the seat. Remote hourly contract at $45–60/hour.

    $45 – $60 / HourOpen to United States
  • Author AI evaluation tasks in budgeting, forecasting, variance analysis and capital allocation: realistic FP&A scenarios, reference models and decks, and rubrics that reward real planning judgment over template work. For FP&A directors and CFOs with 5+ years. $80–90/hour, remote contract.

    $80 – $90 / HourWorldwide
  • Audit repository-level software engineering benchmark tasks for a frontier AI lab: reference patches, test harnesses, Docker isolation, and signs of answer leakage or reward hacking. For US engineers with 3+ years and real open-source contributor or maintainer history. $70–90/hour.

    $70 – $90 / HourOpen to United States
  • Paid pilot for board-certified, practising physicians with deep therapeutic-area expertise: interpret trial endpoints, judge whether results would change prescribing and real-world uptake, and write rubrics that evaluate AI analysis of drugs. 10–20 hours over 1–2 weeks. $150–230/hour.

    $150 – $230 / HourWorldwide
  • Grade AI-generated slides, spreadsheets and documents to consulting standard, flagging factual, visual and presentation errors with structured written feedback. Needs 5+ years at a top firm in the US, UK, Canada, Australia or New Zealand. $100–150/hour.

    $100 – $150 / HourWorldwide
  • Red-team frontier AI models on chemical misuse: write benign, dual-use and adversarial prompts from route design and scale-up experience, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Write and verify expert-level physics multiple-choice questions (one right answer, nine plausible wrong ones) for an AI benchmark, across semiconductors, photonics, quantum sensing, plasma, turbulence and geophysics. PhD or doctoral candidate preferred. Remote hourly contract at $61–77/hour, 10+ hours a week.

    $61 – $77 / HourWorldwide
  • Design Fortune 500 go-to-market scenarios, write reference deal strategies and campaigns, and author rubrics that test whether AI shows real enterprise sales and marketing judgment. For senior F500 sellers, marketers and RevOps leads. Remote hourly contract at $60–70/hour.

    $60 – $70 / HourWorldwide
  • Licensed blasters and blasting engineers red-team frontier AI models: write benign, dual-use and adversarial prompts from quarry, mine and demolition practice, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Senior electrical engineers from large technology, industrial or energy companies build AI evaluation scenarios across power systems, PCB and semiconductor design, embedded firmware, signal processing and controls, with reference analyses and rubrics on a US (NEC, NESC, IEEE) or IEC track. Remote hourly contract at $70–80/hour.

    $70 – $80 / HourWorldwide
  • Civil rights, environmental and legal aid attorneys build AI evaluation scenarios, reference briefs and rubrics on a US track (civil rights statutes, NEPA, CAA, CWA) or an international one (human rights law, EU directives). Remote hourly contract at $90–100/hour; 5+ years in public interest practice.

    $90 – $100 / HourWorldwide
  • A full-time W-2 placement at a leading AI lab through Cincinnatus LLC: US-based mechanical engineers vet model outputs, write instruction specs and reference solutions, and build benchmarks. 5+ years in industry and a mechanical engineering degree required. 40 hours a week for an initial 2–3 months, $60–90/hour.

    $60 – $90 / HourOpen to United States
  • Judge AI-enhanced and upscaled video and stills at pixel level for a leading AI lab's GenAI team: artifacts, noise, aliasing, banding, grain, sharpening. For VFX and rendering supervisors, colorists, DPs, lighting artists and high-end photographers with 5+ years. US, 20 hours a week, $60–90/hour.

    $60 – $90 / HourOpen to United States

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.