Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • Probe AI chat models and agents for safety failures in Danish and English, then turn each failure into labelled, reproducible red-team data. For native Danish speakers with a security, adversarial ML or abuse-analysis background. Remote hourly contract, $48–62/hour, weekly pay.

    $48 – $62 / HourWorldwide
  • Write sensitive-topic prompts in Ukrainian and classify prompts and conversations for an AI safety project, flagging adversarial phrasing as you go. Native Ukrainian, business English and a bachelor's (finished or in progress). Part-time remote at $38–42/hour; Eastern Europe preferred, not required.

    $38 – $42 / HourWorldwide
  • Assess AI-narrated audiobooks in French (France) and annotate each failure: wrong or missing liaisons, mispronounced or skipped words, misread numbers and abbreviations, stilted intonation, glitches. For native French speakers who love audiobooks. Remote hourly contract, around 20 hours a week, $15–20/hour.

    $15 – $20 / HourWorldwide
  • Review AI-narrated audiobooks in US / North American English and flag where the synthetic narrator slips: mispronounced names, dropped or added words, numbers and abbreviations read wrongly, flat intonation, glitches. The best-paid version of this project at $20–25/hour. Remote hourly contract, about 20 hours a week.

    $20 – $25 / HourWorldwide
  • Listen to AI-narrated audiobooks in US / North American Spanish and log each narration error: mispronounced, skipped or added words, misread numbers and abbreviations, stiff intonation, audio glitches. For native Spanish-speaking audiobook listeners. Remote hourly contract, about 20 hours a week, $15–20/hour.

    $15 – $20 / HourWorldwide
  • Adversarial testing of AI chat models in Bahasa Indonesia and English: jailbreaks, prompt injection, bias and multi-turn manipulation, recorded as structured red-team data. Native Indonesian plus prior red-teaming or security experience. Remote hourly contract, $17–25/hour, weekly pay.

    $17 – $25 / HourWorldwide
  • Probe AI chat models and agents for safety failures in Vietnamese and English (jailbreaks, prompt injection, bias, multi-turn manipulation) and document each as reproducible data. Native Vietnamese plus prior red-teaming or security experience. Remote hourly contract, $17–25/hour.

    $17 – $25 / HourWorldwide
  • For Flemish speakers living in Belgium: write sensitive-topic prompts in Belgian Dutch, classify conversations and flag adversarial phrasing so AI models handle Belgian usage safely. Belgium residence required; bachelor's (in progress is fine) and business English. Part-time remote at $48–52/hour.

    $48 – $52 / HourOpen to Belgium
  • Audit Hindi speech data for Amazon's Sonic collections: check annotators' transcriptions against the audio with fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native Hindi (India) is a hard requirement; rulebook and rationales are in English. Remote hourly contract at $18/hour.

    $18 / HourWorldwide
  • Probe AI models in Tamil and English for jailbreaks, bias and harmful output, and assess whether their Tamil answers are accurate and appropriate. Evaluation judgment is the core requirement, not prior red-teaming. Remote hourly contract, $16–22/hour, weekly pay; 57 hired this month.

    $16 – $22 / HourWorldwide
  • Native Finnish speakers write prompts on sensitive topics, classify conversations and flag adversarial phrasing so AI models stay safe in Finnish. Needs business English and a bachelor's, finished or in progress. Part-time remote at $48–52/hour; Finland or Western Europe preferred, not required.

    $48 – $52 / HourWorldwide
  • Red-team AI models in Malay and English: jailbreaks, prompt injection, bias and multi-turn manipulation, captured as labelled, reproducible safety data. Native Malay required, along with prior adversarial, security or abuse-analysis experience. Remote hourly contract at $17–25/hour, paid weekly.

    $17 – $25 / HourWorldwide
  • Audit Mandarin Chinese speech data for two Amazon Sonic collections: judge annotators' transcripts against audio using fixed error codes, and fix word-level timestamp alignment. Native Mandarin as spoken in mainland China is a hard requirement; all rules and rationales are in English. Remote hourly contract at $21.50/hour.

    $21.50 / HourWorldwide
  • Audit Korean speech data for Amazon's Sonic project: check transcripts against audio with fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native Korean as spoken in South Korea is a hard requirement; rationales are written in English. Remote hourly contract at $25/hour.

    $25 / HourWorldwide
  • Audit Japanese speech data for Amazon's Sonic collections: verify annotators' transcripts against the audio with fixed error codes and pass/fail calls, and correct word-level timestamp alignment. Native Japanese as spoken in Japan is a hard requirement; rationales are in English. Remote hourly contract at $37.50/hour.

    $37.50 / HourWorldwide
  • Test AI chat models in Malayalam and English for safety failures and judge whether their answers are accurate, complete and appropriate. Native Malayalam plus careful evaluation skills; prior red-teaming is a plus rather than the core requirement. Remote hourly contract at $16–22/hour, paid weekly.

    $16 – $22 / HourWorldwide
  • Native Telugu speakers in India map the layout of real Telugu PDF pages (headings, tables, figures, reading order) and transcribe every text region exactly in Telugu script, including handwriting, to train document AI. Remote hourly contract at $12.68/hour, with peer review of every task.

    $12.68 / HourOpen to India
  • Native Danish speakers write sensitive-topic prompts, classify prompts and conversations, and flag adversarial phrasing to make AI models safer in Danish. Business English and a bachelor's (in progress counts). Part-time remote at $48–52/hour; Denmark or Western Europe preferred, not required.

    $48 – $52 / HourWorldwide
  • Audit French speech data for Amazon's Sonic project: check annotators' transcripts against audio with fixed error codes and pass/fail calls, and correct word-level timestamp alignment. Native French as spoken in France is a hard requirement; rationales are in English. Joint top rate of the Sonic set at $41.50/hour, remote contract.

    $41.50 / HourWorldwide
  • Native Gujarati speakers in India annotate the structure of real Gujarati PDF pages and transcribe every text region exactly in Gujarati script, handwriting included, to build training data for document AI. Remote hourly contract at $12.68/hour; every task is peer-reviewed by a second Gujarati expert.

    $12.68 / HourOpen to India
  • Audit Italian speech data for Amazon's Sonic collections: judge annotators' transcripts against audio with fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native Italian as spoken in Italy is a hard requirement; rationales are written in English. Remote hourly contract at $39.50/hour.

    $39.50 / HourWorldwide
  • Audit German speech data for Amazon's Sonic collections: verify annotators' transcripts against audio using fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native German as spoken in Germany is a hard requirement; rationales are in English. Joint top rate of the Sonic set at $41.50/hour, remote.

    $41.50 / HourWorldwide
  • Native Bengali speakers based in India map the layout of real Bengali PDF pages and transcribe every text region character for character in Bengali script, handwriting included, to train document AI. Remote hourly contract at $12.68/hour, with a second-expert review of every task.

    $12.68 / HourOpen to India
  • Test AI chat models in Marathi and English for safety failures (jailbreaks, bias, harmful answers) and judge whether their Marathi is accurate and appropriate rather than Hindi in disguise. Evaluation judgment is the core ask. Remote hourly contract, $16–22/hour, weekly pay.

    $16 – $22 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.