Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • The best-paid role in Mercor's bilingual AI safety series: Norwegian-speaking PhD chemists and biologists write specialised science prompts and grade how AI models handle dual-use questions. Part-time remote at $77–81/hour. Norway preferred, not required; PhD candidates eligible.

    $77 – $81 / HourWorldwide
  • Adversarial testing of AI chat models in Bahasa Indonesia and English: jailbreaks, prompt injection, bias and multi-turn manipulation, recorded as structured red-team data. Native Indonesian plus prior red-teaming or security experience. Remote hourly contract, $17–25/hour, weekly pay.

    $17 – $25 / HourWorldwide
  • Produce or judge expert finance deliverables for AI: PnL and AP reconciliations, 3-statement and pro-forma models with purchase accounting, covenant analysis and partnership allocations, all from messy source documents. For product control, FP&A, technical accounting and treasury people with 3+ years. $70/hour, remote.

    $70 / HourWorldwide
  • Solve and evaluate complex electrical design problems (power distribution, protection coordination, equipment sizing, load flow) on a short-term expert evaluation project for a leading technology organization. For practising electrical design engineers. Remote, $70–85/hour.

    $70 – $85 / HourWorldwide
  • Probe AI chat models and agents for safety failures in Vietnamese and English (jailbreaks, prompt injection, bias, multi-turn manipulation) and document each as reproducible data. Native Vietnamese plus prior red-teaming or security experience. Remote hourly contract, $17–25/hour.

    $17 – $25 / HourWorldwide
  • Remote hourly contract for power users, IT administrators and desktop support staff who run Windows 11 Pro as their daily OS. $60–70/hour, paid weekly via Stripe or Wise. Home edition users may not qualify, and a single 1080p screen does not meet the display rule. Tasks are not described.

    $60 – $70 / HourWorldwide
  • Run 20-minute structured interviews that vet senior engineers (GPU kernels, security research, ML compilers) for a frontier lab's AI training and evaluation panel, then tier each candidate and write a summary. For experienced recruiters and technical vetters. US, part-time, about 10 hours a week, $50–60/hour.

    $50 – $60 / HourOpen to United States
  • Croatian-speaking PhD chemists and biologists write specialised science prompts in Croatian and evaluate how AI models answer, including how they handle dual-use questions. Part-time remote at $48–52/hour, $10 above the Croatian generalist role. Eastern Europe preferred, not required; 9 hires this month.

    $48 – $52 / HourWorldwide
  • Red-team AI models in Malay and English: jailbreaks, prompt injection, bias and multi-turn manipulation, captured as labelled, reproducible safety data. Native Malay required, along with prior adversarial, security or abuse-analysis experience. Remote hourly contract at $17–25/hour, paid weekly.

    $17 – $25 / HourWorldwide
  • Remote hourly contract for lawyers, policy staff, consultants and technical writers who author long documents in Word on their own Mac. $60–70/hour, paid weekly via Stripe or Wise. Requires your own Microsoft licence and a screen above 2.5 megapixels. Tasks are not described.

    $60 – $70 / HourWorldwide
  • Remote hourly contract for consultants, analysts and marketers who build decks in PowerPoint on their own Windows PC. $60–70/hour, paid weekly via Stripe or Wise. Your own Microsoft licence and a screen setup above 2.5 megapixels are required; the ad does not describe the tasks.

    $60 – $70 / HourWorldwide
  • QA engineers, SDETs and test automation engineers review browser-based test workflows for AI-generated web apps, checking that each test is feasible, isolated, deterministic and gives a clean pass or fail before it joins a benchmark dataset. 3+ years and Playwright, Cypress or Selenium experience. Remote hourly contract at $30–60/hour.

    $30 – $60 / HourWorldwide
  • Full-time W-2 coordinator (via Cincinnatus LLC) placed with a leading AI lab's GenAI team, running day-to-day operations on AI training-data projects: turning leadership input into expert guidelines, onboarding experts, tracking quality and deliverables. US-based, 3+ years of coordination plus hands-on AI data experience. $45–55/hour.

    $45 – $55 / HourOpen to United States
  • Pediatric acute care RNs evaluate AI outputs built from nursing flowsheet documentation, annotate pediatric assessment data and help write clinical benchmarks. US licence outside California, inpatient bedside work within the past 10 years. $55–65/hour.

    $55 – $65 / HourOpen to United States
  • Belgium-based PhD chemists and biologists who write Belgian Dutch: author specialised science prompts and grade how AI models handle accuracy and dual-use safety. Belgium residence required. Part-time remote at $61–65/hour, $13 above the Belgian Dutch generalist role; 9 hires this month.

    $61 – $65 / HourOpen to Belgium
  • Inpatient registered nurses with broad clinical exposure support ongoing clinical AI product work: annotation, clinical review, model evaluation, user-feedback investigation and guideline development, alongside engineers and clinicians. US only, 10 hours a week minimum, $55–65/hour.

    $55 – $65 / HourOpen to United States
  • Design finance Excel tasks from your own work (three-statement models, LBO and DCF valuation, forecasting), write model solutions, and evaluate AI attempts for a leading tech company's GenAI team. US only, W-2 through Cincinnatus LLC, 40 hours a week to end of September then 20. $70–100/hour.

    $70 – $100 / HourOpen to United States
  • Audit end-to-end AI-assisted coding sessions (traces from tools like Cursor, Copilot or Claude Code) used to train and evaluate a frontier lab's models, judging correctness, workflow and reasoning with rubric-based feedback. 3+ years of software development plus hands-on agentic coding. US-only, $70–90/hour.

    $70 – $90 / HourOpen to United States
  • ML systems engineers write and evaluate training tasks for a frontier lab across GPU kernels, performance profiling, distributed debugging and LLM inference serving, plus the rubrics that grade them. 2+ years of hands-on ML infrastructure work. Canada, UK or US; 40 hours a week. $90–120/hour.

    $90 – $120 / HourOpen to Canada, United Kingdom and 1 more country
  • Audit applied machine-learning tasks used to train and evaluate a frontier AI lab's models: experiment design, model selection, evaluation methodology, leakage and metric gaming. For US practitioners with 3+ years of hands-on experimental ML in PyTorch, TensorFlow, scikit-learn or XGBoost. $70–90/hour.

    $70 – $90 / HourOpen to United States
  • Design Excel tasks from cross-industry experience (data models, dashboards, Power Query, VBA automation), write the solutions, and evaluate AI attempts for a tech company's GenAI team. Breadth across functions, teaching or a PhD help. US only, W-2 via Cincinnatus LLC, 40 then 20 hours a week. $70–100/hour.

    $70 – $100 / HourOpen to United States
  • Korean-speaking PhD chemists and biologists write specialised science prompts in Korean and grade AI answers for accuracy and dual-use safety. Part-time remote at $63–67/hour, $15 above the Korean generalist role and third-highest in the series. East Asia preferred, not required.

    $63 – $67 / HourWorldwide
  • Hindi-speaking PhD chemists and biologists write specialised science prompts in Hindi and grade AI answers for accuracy and safe handling of dual-use topics. Part-time remote at $23–27/hour; India or South Asia preferred, not required. 12 hires this month, among the busiest in the series.

    $23 – $27 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.