Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • The highest-paid generalist role in Mercor's bilingual AI safety series: native Norwegian speakers write sensitive-topic prompts, classify conversations and flag adversarial phrasing. Bachelor's (in progress is fine) and business English. Part-time remote at $58–62/hour; Norway preferred, not required.

    $58 – $62 / HourWorldwide
  • Materials science PhDs author original, executable research problems for a scientific-computing AI benchmark, with depth in both semiconductor materials and molecular modeling. Tasks ship only when frontier models fail them more often than not. 6 weeks, 20+ hours a week, Git and Docker workflow. $70/hour, 212 hired this month.

    $70 / HourWorldwide
  • Ukrainian-speaking PhD chemists and biologists write specialised science prompts in Ukrainian and grade AI answers for accuracy and safe handling of dual-use topics. Part-time remote at $48–52/hour, $10 above the Ukrainian generalist role. Eastern Europe preferred, not required; 11 hires this month.

    $48 – $52 / HourWorldwide
  • Radiological emergency planners, field monitoring teams and consequence modellers red-team frontier AI models: write benign, dual-use and adversarial prompts, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Audit Brazilian Portuguese speech data for Amazon's Sonic project: check transcripts against audio with fixed error codes and pass/fail verdicts, and fix word-level timestamp alignment. Native Portuguese as spoken in Brazil is a hard requirement; rationales are written in English. Remote hourly contract at $25/hour.

    $25 / HourWorldwide
  • Quality-check AI-narrated audiobooks in German (Germany): pinpoint mispronounced, missing or extra words, wrongly read numbers and abbreviations, unnatural prosody and audio artefacts, and judge the overall listen. For native German audiobook listeners. Remote hourly contract, about 20 hours a week, $15–20/hour.

    $15 – $20 / HourWorldwide
  • Japanese-speaking PhD chemists and biologists write specialised science prompts in Japanese and grade AI answers for accuracy and dual-use safety. Part-time remote at $68–72/hour, second-highest in the series and $20 above the Japanese generalist role. East Asia preferred, not required.

    $68 – $72 / HourWorldwide
  • Remote hourly contract for FPGA, digital design and embedded hardware engineers who already run Intel/Altera Quartus II on their own Windows PC. $55–65/hour, paid weekly via Stripe or Wise, $10 more than the sister Vivado listing. Needs a display above 2.5 megapixels.

    $55 – $65 / HourWorldwide
  • Native or near-native English editors, proofreaders and linguists evaluate LLM-generated English, rewrite it to a high stylistic standard, annotate grammatical and stylistic features, and give feedback that guides model training at a leading AI lab. Remote, flexible hours, $50/hour.

    $50 / HourWorldwide
  • A paid expert conversation for engineers who build and run LLM agents in production: a short AI screening interview (no coding), then, if selected, a 30-minute live call on agent reliability, evaluation and internal adoption, paid $100–500 depending on depth of experience.

    $100 – $500 / TaskWorldwide
  • Paid pilot for a biotech and pharma research team: create and critique rubrics that assess commercial drugs and development programs, judge investment-style theses, and assess 5–10 companies end to end. For specialist-fund biotech analysts with 5+ years. About 10–20 hours over 1–2 weeks, $120–200/hour.

    $120 – $200 / HourWorldwide
  • Red-team frontier AI models on chemical safety: write benign, dual-use and adversarial prompts from exposure and process-hazard work, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Umbrella listing for Mercor's chemistry red-team panel: synthetic, analytical, forensic, defence and process safety chemists write benign, dual-use and adversarial prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 5 hired this month.

    $65 – $75 / TaskWorldwide
  • Red-team frontier AI models on chemical misuse: write benign, dual-use and adversarial prompts from trace analysis and method validation, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Remote hourly contract for FPGA and RTL engineers who already have AMD/Xilinx Vivado running on their own Windows PC. $45–55/hour, paid weekly via Stripe or Wise. You supply the tool, the hardware and a screen above 2.5 megapixels; the tasks themselves are not described.

    $45 – $55 / HourWorldwide
  • Remote hourly contract for music producers, beatmakers and audio engineers who already produce in FL Studio (FruityLoops) on their own Windows PC. $30–40/hour, paid weekly via Stripe or Wise. You need your own copy and a display above 2.5 megapixels; tasks are not described.

    $30 – $40 / HourWorldwide
  • Judge AI-generated Thai song lyrics for a leading AI lab: compare them with published songs for similarity, rate quality, creativity, prompt adherence and originality, and check that slang and word choice sound natural. For Thai songwriters, lyricists, performers or music journalists. Remote, flexible, up to 6 months, $18/hour.

    $18 / HourWorldwide
  • Evaluate AI-narrated audiobooks in Dutch (Netherlands): mark each mispronounced, skipped or added word, misread number, awkward intonation and audio glitch, then rate the overall listen. For native Dutch speakers who are regular audiobook listeners. Remote hourly contract, about 20 hours a week, $15–20/hour.

    $15 – $20 / HourWorldwide
  • Listen to AI-narrated audiobooks in Brazilian Portuguese and tag where the synthetic narrator slips: mispronunciations, skipped or extra words, wrong readings of numbers and abbreviations, flat or odd intonation. For native Brazilian audiobook listeners. Remote hourly contract, around 20 hours a week, $15–20/hour.

    $15 – $20 / HourWorldwide
  • Author AI evaluation tasks in budgeting, forecasting, variance analysis and capital allocation: realistic FP&A scenarios, reference models and decks, and rubrics that reward real planning judgment over template work. For FP&A directors and CFOs with 5+ years. $80–90/hour, remote contract.

    $80 – $90 / HourWorldwide
  • Test AI models for safety failures in Finnish and English: jailbreaks, prompt injection, bias and multi-turn manipulation, logged as structured red-team data. For native Finnish speakers with prior adversarial, security or abuse-analysis experience. Remote hourly contract, $48–62/hour.

    $48 – $62 / HourWorldwide
  • Swedish and English red-teaming of AI models: jailbreaks, injected instructions, bias and multi-turn manipulation, written up as reproducible attack cases. The busiest listing in this family, with 243 hires this month. Remote hourly contract at $48–62/hour, paid weekly.

    $48 – $62 / HourWorldwide
  • Remote hourly contract for mechanical designers and drafters who already run AutoCAD Mechanical on their own Windows PC for 2D drafting and detailing. $45–55/hour, paid weekly via Stripe or Wise. Plain AutoCAD may not be enough; the ad names the Mechanical toolset.

    $45 – $55 / HourWorldwide
  • Test AI models in Gujarati and English for jailbreaks, bias and harmful answers, and judge whether their Gujarati is accurate and appropriate. Evaluation judgment is the core requirement. Open to the Gujarati diaspora, with no residence rule published. Remote hourly contract at $16–22/hour; 56 hired this month.

    $16 – $22 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.