AI safety testing in Telugu and English: provoke and document jailbreaks, bias and harmful output from chat models, and judge the quality of their Telugu answers. Evaluation skill matters more than security experience. Remote hourly contract at $16–22/hour, weekly pay; H-1B and STEM OPT excluded.
Remote AI training and data labeling jobs
Filter jobs
Location
- Worldwide8 jobs
- United Statesno roles alongside your other filters
- United Kingdomno roles alongside your other filters
- Canadano roles alongside your other filters
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
- English8 jobs
- Germanno roles alongside your other filters
- Spanishno roles alongside your other filters
- Frenchno roles alongside your other filters
- Japaneseno roles alongside your other filters
- Portugueseno roles alongside your other filters
Field: AI Safety & Evaluation
- Languages & Linguistics82 jobs
- Audio & Voice65 jobs
- Engineering4 jobs
- Business & Finance9 jobs
- Software & IT11 jobs
- Health & Medicine5 jobs
- Law, Policy & Security3 jobs
- General & Data Collection12 jobs
- Science & Math1 job
- AI Safety & Evaluation8 jobs, applied. Activate to remove
- Video, Image & Design13 jobs
- Data, AI & ML3 jobs
- Writing & Education4 jobs
- Other fieldsno roles alongside your other filters
Newest
8 open roles matching these filters
- $16 – $22 / HourWorldwide
Odia and English AI safety work: probe chat models for jailbreaks, bias and harmful output and judge whether their Odia answers hold up. One of the lowest-resource languages in the family and the busiest Indian variant, with 88 hired this month. Remote hourly contract, $16–22/hour.
$16 – $22 / HourWorldwideAI safety work in Kannada and English: test chat models for jailbreaks, bias and harmful output, and judge whether their Kannada answers are accurate and appropriate. Unlike the European variants, prior red-teaming is not listed as a requirement. Remote hourly contract, $16–22/hour, weekly pay.
$16 – $22 / HourWorldwideTest AI models in Gujarati and English for jailbreaks, bias and harmful answers, and judge whether their Gujarati is accurate and appropriate. Evaluation judgment is the core requirement. Open to the Gujarati diaspora, with no residence rule published. Remote hourly contract at $16–22/hour; 56 hired this month.
$16 – $22 / HourWorldwideProbe AI models in Tamil and English for jailbreaks, bias and harmful output, and assess whether their Tamil answers are accurate and appropriate. Evaluation judgment is the core requirement, not prior red-teaming. Remote hourly contract, $16–22/hour, weekly pay; 57 hired this month.
$16 – $22 / HourWorldwideTest AI chat models in Malayalam and English for safety failures and judge whether their answers are accurate, complete and appropriate. Native Malayalam plus careful evaluation skills; prior red-teaming is a plus rather than the core requirement. Remote hourly contract at $16–22/hour, paid weekly.
$16 – $22 / HourWorldwideTest AI chat models in Marathi and English for safety failures (jailbreaks, bias, harmful answers) and judge whether their Marathi is accurate and appropriate rather than Hindi in disguise. Evaluation judgment is the core ask. Remote hourly contract, $16–22/hour, weekly pay.
$16 – $22 / HourWorldwideAI safety testing in Urdu and English: provoke jailbreaks, bias and harmful output from chat models and judge whether their Urdu answers are accurate and appropriate, in Nastaliq script or Roman Urdu. Evaluation judgment is the core ask. Remote hourly contract at $16–22/hour, weekly pay.
$16 – $22 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.