Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • Build AI evaluation tasks set inside Fortune 500 insurance: commercial underwriting, claims adjudication, reserving, reinsurance and NAIC compliance scenarios, with reference guidelines and rubrics. Needs 5+ years at a major carrier or reinsurer. $50–60/hour, remote contract.

    $50 – $60 / HourWorldwide
  • The highest-paid generalist role in Mercor's bilingual AI safety series: native Norwegian speakers write sensitive-topic prompts, classify conversations and flag adversarial phrasing. Bachelor's (in progress is fine) and business English. Part-time remote at $58–62/hour; Norway preferred, not required.

    $58 – $62 / HourWorldwide
  • Materials science PhDs author original, executable research problems for a scientific-computing AI benchmark, with depth in both semiconductor materials and molecular modeling. Tasks ship only when frontier models fail them more often than not. 6 weeks, 20+ hours a week, Git and Docker workflow. $70/hour, 212 hired this month.

    $70 / HourWorldwide
  • The top tier of Mercor's embedded legal expert role: a full-time W-2 job (via Cincinnatus LLC) with a leading AI lab in the Bay Area, reviewing legal model outputs, writing golden solutions and building benchmarks. For partners and general counsel. Hybrid, 6-month initial term, $100–150/hour.

    $100 – $150 / HourHybridOpen to United States
  • Ukrainian-speaking PhD chemists and biologists write specialised science prompts in Ukrainian and grade AI answers for accuracy and safe handling of dual-use topics. Part-time remote at $48–52/hour, $10 above the Ukrainian generalist role. Eastern Europe preferred, not required; 11 hires this month.

    $48 – $52 / HourWorldwide
  • Germany-based native German speakers record one session of about four hours, which a Mercor client uses to clone the voice for its internal customer-experience AI agent. $50–100/hour in USD, no acting experience required. You are licensing your voice; the ad limits its use but leaves duration and withdrawal open.

    $50 – $100 / HourOpen to Germany
  • Radiological emergency planners, field monitoring teams and consequence modellers red-team frontier AI models: write benign, dual-use and adversarial prompts, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • A hybrid, Bay Area-based W-2 role embedded with a leading AI lab: senior software engineers vet model outputs, write instruction specs and golden solutions, and build engineering benchmarks. 4+ years, senior-level progression and a CS or engineering degree. 40 hours a week for an initial 6 months, $65–105/hour.

    $65 – $105 / HourHybridOpen to United States
  • Audit Brazilian Portuguese speech data for Amazon's Sonic project: check transcripts against audio with fixed error codes and pass/fail verdicts, and fix word-level timestamp alignment. Native Portuguese as spoken in Brazil is a hard requirement; rationales are written in English. Remote hourly contract at $25/hour.

    $25 / HourWorldwide
  • Quality-check AI-narrated audiobooks in German (Germany): pinpoint mispronounced, missing or extra words, wrongly read numbers and abbreviations, unnatural prosody and audio artefacts, and judge the overall listen. For native German audiobook listeners. Remote hourly contract, about 20 hours a week, $15–20/hour.

    $15 – $20 / HourWorldwide
  • Native Odiya (Oriya) speakers in India annotate the layout of real Odiya PDF pages and transcribe each text region exactly in Odiya script, handwriting included, to build training data for document AI. Remote hourly contract at $12.68/hour, with every task reviewed by a second Odiya expert.

    $12.68 / HourOpen to India
  • Japanese-speaking PhD chemists and biologists write specialised science prompts in Japanese and grade AI answers for accuracy and dual-use safety. Part-time remote at $68–72/hour, second-highest in the series and $20 above the Japanese generalist role. East Asia preferred, not required.

    $68 – $72 / HourWorldwide
  • Remote hourly contract for FPGA, digital design and embedded hardware engineers who already run Intel/Altera Quartus II on their own Windows PC. $55–65/hour, paid weekly via Stripe or Wise, $10 more than the sister Vivado listing. Needs a display above 2.5 megapixels.

    $55 – $65 / HourWorldwide
  • Native or near-native English editors, proofreaders and linguists evaluate LLM-generated English, rewrite it to a high stylistic standard, annotate grammatical and stylistic features, and give feedback that guides model training at a leading AI lab. Remote, flexible hours, $50/hour.

    $50 / HourWorldwide
  • A paid expert conversation for engineers who build and run LLM agents in production: a short AI screening interview (no coding), then, if selected, a 30-minute live call on agent reliability, evaluation and internal adoption, paid $100–500 depending on depth of experience.

    $100 – $500 / TaskWorldwide
  • Paid pilot for a biotech and pharma research team: create and critique rubrics that assess commercial drugs and development programs, judge investment-style theses, and assess 5–10 companies end to end. For specialist-fund biotech analysts with 5+ years. About 10–20 hours over 1–2 weeks, $120–200/hour.

    $120 – $200 / HourWorldwide
  • Red-team frontier AI models on chemical safety: write benign, dual-use and adversarial prompts from exposure and process-hazard work, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Full-time Bay Area hybrid role with an AI lab's research team: review model reasoning on life sciences tasks, write golden solutions and instruction specs, and design benchmarks. For life sciences PhDs with 4+ years of substantive research experience. W-2 via Cincinnatus, $65–105/hour.

    $65 – $105 / HourHybridOpen to United States
  • Umbrella listing for Mercor's chemistry red-team panel: synthetic, analytical, forensic, defence and process safety chemists write benign, dual-use and adversarial prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 5 hired this month.

    $65 – $75 / TaskWorldwide
  • Native Malayalam speakers in India segment real Malayalam PDF pages into typed, ordered regions and transcribe every one exactly in Malayalam script, handwriting included, for document AI training data. Remote hourly contract at $12.68/hour; each task gets a full second-expert review.

    $12.68 / HourOpen to India
  • Red-team frontier AI models on chemical misuse: write benign, dual-use and adversarial prompts from trace analysis and method validation, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.

    $65 – $75 / TaskWorldwide
  • Remote hourly contract for FPGA and RTL engineers who already have AMD/Xilinx Vivado running on their own Windows PC. $45–55/hour, paid weekly via Stripe or Wise. You supply the tool, the hardware and a screen above 2.5 megapixels; the tasks themselves are not described.

    $45 – $55 / HourWorldwide
  • Work mock AML, KYC, sanctions and fraud cases end to end and produce the analyst deliverable (alert dispositions, SAR narratives, remediation lists), graded against a rubric to build AI training data. US-based, about 15 hours a week, $75–100/hour. Needs 3+ years in financial crime.

    $75 – $100 / HourOpen to United States
  • Remote hourly contract for music producers, beatmakers and audio engineers who already produce in FL Studio (FruityLoops) on their own Windows PC. $30–40/hour, paid weekly via Stripe or Wise. You need your own copy and a display above 2.5 megapixels; tasks are not described.

    $30 – $40 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.