Probe AI chat models and agents for safety failures in Danish and English, then turn each failure into labelled, reproducible red-team data. For native Danish speakers with a security, adversarial ML or abuse-analysis background. Remote hourly contract, $48–62/hour, weekly pay.
Remote AI training and data labeling jobs
Filter jobs
Location
- Worldwide50 jobs
- United States2 jobs
- United Kingdom1 job
- Canada1 job
- India5 jobs
- Mexicono roles alongside your other filters
Language
Field: Languages & Linguistics
- Languages & Linguistics58 jobs, applied. Activate to remove
- Audio & Voice43 jobs
- Engineering40 jobs
- Business & Finance52 jobs
- Software & IT32 jobs
- Health & Medicine28 jobs
- Law, Policy & Security17 jobs
- General & Data Collection42 jobs
- Science & Math41 jobs
- AI Safety & Evaluation61 jobs
- Video, Image & Design9 jobs
- Data, AI & ML15 jobs
- Writing & Education1 job
- Other fieldsno roles alongside your other filters
Level
- Entry12 jobs
- Junior31 jobs
- Medium15 jobs
- Seniorno roles alongside your other filters
Newest
58 open roles matching these filters · page 2 of 3
- $48 – $62 / HourWorldwide
Write sensitive-topic prompts in Ukrainian and classify prompts and conversations for an AI safety project, flagging adversarial phrasing as you go. Native Ukrainian, business English and a bachelor's (finished or in progress). Part-time remote at $38–42/hour; Eastern Europe preferred, not required.
$38 – $42 / HourWorldwideAssess AI-narrated audiobooks in French (France) and annotate each failure: wrong or missing liaisons, mispronounced or skipped words, misread numbers and abbreviations, stilted intonation, glitches. For native French speakers who love audiobooks. Remote hourly contract, around 20 hours a week, $15–20/hour.
$15 – $20 / HourWorldwideReview AI-narrated audiobooks in US / North American English and flag where the synthetic narrator slips: mispronounced names, dropped or added words, numbers and abbreviations read wrongly, flat intonation, glitches. The best-paid version of this project at $20–25/hour. Remote hourly contract, about 20 hours a week.
$20 – $25 / HourWorldwideListen to AI-narrated audiobooks in US / North American Spanish and log each narration error: mispronounced, skipped or added words, misread numbers and abbreviations, stiff intonation, audio glitches. For native Spanish-speaking audiobook listeners. Remote hourly contract, about 20 hours a week, $15–20/hour.
$15 – $20 / HourWorldwideAdversarial testing of AI chat models in Bahasa Indonesia and English: jailbreaks, prompt injection, bias and multi-turn manipulation, recorded as structured red-team data. Native Indonesian plus prior red-teaming or security experience. Remote hourly contract, $17–25/hour, weekly pay.
$17 – $25 / HourWorldwideProbe AI chat models and agents for safety failures in Vietnamese and English (jailbreaks, prompt injection, bias, multi-turn manipulation) and document each as reproducible data. Native Vietnamese plus prior red-teaming or security experience. Remote hourly contract, $17–25/hour.
$17 – $25 / HourWorldwideFor Flemish speakers living in Belgium: write sensitive-topic prompts in Belgian Dutch, classify conversations and flag adversarial phrasing so AI models handle Belgian usage safely. Belgium residence required; bachelor's (in progress is fine) and business English. Part-time remote at $48–52/hour.
$48 – $52 / HourOpen to BelgiumAudit Hindi speech data for Amazon's Sonic collections: check annotators' transcriptions against the audio with fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native Hindi (India) is a hard requirement; rulebook and rationales are in English. Remote hourly contract at $18/hour.
$18 / HourWorldwideProbe AI models in Tamil and English for jailbreaks, bias and harmful output, and assess whether their Tamil answers are accurate and appropriate. Evaluation judgment is the core requirement, not prior red-teaming. Remote hourly contract, $16–22/hour, weekly pay; 57 hired this month.
$16 – $22 / HourWorldwideNative Finnish speakers write prompts on sensitive topics, classify conversations and flag adversarial phrasing so AI models stay safe in Finnish. Needs business English and a bachelor's, finished or in progress. Part-time remote at $48–52/hour; Finland or Western Europe preferred, not required.
$48 – $52 / HourWorldwideRed-team AI models in Malay and English: jailbreaks, prompt injection, bias and multi-turn manipulation, captured as labelled, reproducible safety data. Native Malay required, along with prior adversarial, security or abuse-analysis experience. Remote hourly contract at $17–25/hour, paid weekly.
$17 – $25 / HourWorldwideAudit Mandarin Chinese speech data for two Amazon Sonic collections: judge annotators' transcripts against audio using fixed error codes, and fix word-level timestamp alignment. Native Mandarin as spoken in mainland China is a hard requirement; all rules and rationales are in English. Remote hourly contract at $21.50/hour.
$21.50 / HourWorldwideAudit Korean speech data for Amazon's Sonic project: check transcripts against audio with fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native Korean as spoken in South Korea is a hard requirement; rationales are written in English. Remote hourly contract at $25/hour.
$25 / HourWorldwideAudit Japanese speech data for Amazon's Sonic collections: verify annotators' transcripts against the audio with fixed error codes and pass/fail calls, and correct word-level timestamp alignment. Native Japanese as spoken in Japan is a hard requirement; rationales are in English. Remote hourly contract at $37.50/hour.
$37.50 / HourWorldwideTest AI chat models in Malayalam and English for safety failures and judge whether their answers are accurate, complete and appropriate. Native Malayalam plus careful evaluation skills; prior red-teaming is a plus rather than the core requirement. Remote hourly contract at $16–22/hour, paid weekly.
$16 – $22 / HourWorldwideNative Telugu speakers in India map the layout of real Telugu PDF pages (headings, tables, figures, reading order) and transcribe every text region exactly in Telugu script, including handwriting, to train document AI. Remote hourly contract at $12.68/hour, with peer review of every task.
$12.68 / HourOpen to IndiaNative Danish speakers write sensitive-topic prompts, classify prompts and conversations, and flag adversarial phrasing to make AI models safer in Danish. Business English and a bachelor's (in progress counts). Part-time remote at $48–52/hour; Denmark or Western Europe preferred, not required.
$48 – $52 / HourWorldwideAudit French speech data for Amazon's Sonic project: check annotators' transcripts against audio with fixed error codes and pass/fail calls, and correct word-level timestamp alignment. Native French as spoken in France is a hard requirement; rationales are in English. Joint top rate of the Sonic set at $41.50/hour, remote contract.
$41.50 / HourWorldwideNative Gujarati speakers in India annotate the structure of real Gujarati PDF pages and transcribe every text region exactly in Gujarati script, handwriting included, to build training data for document AI. Remote hourly contract at $12.68/hour; every task is peer-reviewed by a second Gujarati expert.
$12.68 / HourOpen to IndiaAudit Italian speech data for Amazon's Sonic collections: judge annotators' transcripts against audio with fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native Italian as spoken in Italy is a hard requirement; rationales are written in English. Remote hourly contract at $39.50/hour.
$39.50 / HourWorldwideAudit German speech data for Amazon's Sonic collections: verify annotators' transcripts against audio using fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native German as spoken in Germany is a hard requirement; rationales are in English. Joint top rate of the Sonic set at $41.50/hour, remote.
$41.50 / HourWorldwideNative Bengali speakers based in India map the layout of real Bengali PDF pages and transcribe every text region character for character in Bengali script, handwriting included, to train document AI. Remote hourly contract at $12.68/hour, with a second-expert review of every task.
$12.68 / HourOpen to IndiaTest AI chat models in Marathi and English for safety failures (jailbreaks, bias, harmful answers) and judge whether their Marathi is accurate and appropriate rather than Hindi in disguise. Evaluation judgment is the core ask. Remote hourly contract, $16–22/hour, weekly pay.
$16 – $22 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.