AI safety testing in Telugu and English: provoke and document jailbreaks, bias and harmful output from chat models, and judge the quality of their Telugu answers. Evaluation skill matters more than security experience. Remote hourly contract at $16–22/hour, weekly pay; H-1B and STEM OPT excluded.
Remote AI training and data labeling jobs
Filter jobs
Location
- Worldwide57 jobs
- United States2 jobs
- United Kingdom2 jobs
- Canadano roles alongside your other filters
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
- Australiano roles alongside your other filters
- Belgium4 jobs
- Ireland2 jobs
- Argentinano roles alongside your other filters
- Switzerland2 jobs
- Chileno roles alongside your other filters
- Colombiano roles alongside your other filters
- Costa Ricano roles alongside your other filters
- Germany2 jobs
- Spain2 jobs
- Panamano roles alongside your other filters
- Brazilno roles alongside your other filters
- Dominican Republicno roles alongside your other filters
- France2 jobs
- Italy2 jobs
- Luxembourg2 jobs
- Netherlands2 jobs
- New Zealandno roles alongside your other filters
- Peruno roles alongside your other filters
- Uruguayno roles alongside your other filters
- Albania2 jobs
- Austria2 jobs
- Bosnia & Herzegovina2 jobs
- Barbadosno roles alongside your other filters
- Bulgaria2 jobs
- Bahamasno roles alongside your other filters
- Czechia2 jobs
- Denmark2 jobs
- Ecuadorno roles alongside your other filters
- Estonia2 jobs
- Finland2 jobs
- Greece2 jobs
- Guatemalano roles alongside your other filters
- Hondurasno roles alongside your other filters
- Croatia2 jobs
- Hungary2 jobs
- Iceland2 jobs
- Jamaicano roles alongside your other filters
- South Koreano roles alongside your other filters
- Liechtenstein2 jobs
- Lithuania2 jobs
- Latvia2 jobs
- Monaco2 jobs
- Moldova2 jobs
- North Macedonia2 jobs
- Malta2 jobs
- Nicaraguano roles alongside your other filters
- Norway2 jobs
- Poland2 jobs
- Portugal2 jobs
- Romania2 jobs
- Serbia2 jobs
- Sweden2 jobs
- Slovenia2 jobs
- Slovakia2 jobs
- San Marino2 jobs
- El Salvadorno roles alongside your other filters
- Kosovo2 jobs
- Boliviano roles alongside your other filters
- Belizeno roles alongside your other filters
- Cubano roles alongside your other filters
- Indonesiano roles alongside your other filters
- Japanno roles alongside your other filters
- Paraguayno roles alongside your other filters
- Venezuelano roles alongside your other filters
- South Africano roles alongside your other filters
Language
- English61 jobs
- German1 job
- Spanishno roles alongside your other filters
- French1 job
- Japanese2 jobs
- Portuguese2 jobs
Field: AI Safety & Evaluation
- Languages & Linguistics58 jobs
- Audio & Voice43 jobs
- Engineering40 jobs
- Business & Finance52 jobs
- Software & IT32 jobs
- Health & Medicine28 jobs
- Law, Policy & Security17 jobs
- General & Data Collection42 jobs
- Science & Math41 jobs
- AI Safety & Evaluation61 jobs, applied. Activate to remove
- Video, Image & Design9 jobs
- Data, AI & ML15 jobs
- Writing & Education1 job
- Other fieldsno roles alongside your other filters
Newest
61 open roles matching these filters · page 1 of 3
- $16 – $22 / HourWorldwide
Odia and English AI safety work: probe chat models for jailbreaks, bias and harmful output and judge whether their Odia answers hold up. One of the lowest-resource languages in the family and the busiest Indian variant, with 88 hired this month. Remote hourly contract, $16–22/hour.
$16 – $22 / HourWorldwideRed-team AI models in European and other non-Brazilian Portuguese plus English: jailbreaks, prompt injection, bias and manipulation, logged as reproducible safety data. Brazilian Portuguese is explicitly excluded. Remote hourly contract at $29–45/hour, paid weekly.
$29 – $45 / HourWorldwideEvaluate frontier AI responses on grey-area and policy-sensitive topics (misinformation, political persuasion, self-harm, violence, cyber, biosecurity), apply safety rubrics and write structured feedback. 5+ years in trust and safety, journalism, policy, research or security. US, UK and most of Europe. $60–70/hour.
$60 – $70 / HourOpen to Albania, Austria and 38 more countriesNative Japanese speakers write prompts on sensitive subjects, classify conversations and flag adversarial phrasing so AI models behave safely in Japanese. Business English and a bachelor's (in progress is fine). Part-time remote at $48–52/hour; Japan or East Asia preferred, not required; 8 hires this month.
$48 – $52 / HourWorldwideCertified explosives specialists, forensic analysts and licensee inspectors red-team frontier AI models: write benign, dual-use and adversarial prompts from casework, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideRed-team conversational AI models in Norwegian and English: attempt jailbreaks, prompt injections and multi-turn manipulation, then document what broke. Prior red-teaming, security or adversarial testing experience expected. Remote hourly contract at $48–62/hour, paid weekly.
$48 – $62 / HourWorldwideNative Croatian speakers write prompts on sensitive subjects, classify prompts and conversations, and flag adversarial phrasing to harden AI models in Croatian. Bachelor's (in progress is fine) and business English. Part-time remote at $38–42/hour; Croatia or Eastern Europe preferred, not required.
$38 – $42 / HourWorldwideAI safety work in Kannada and English: test chat models for jailbreaks, bias and harmful output, and judge whether their Kannada answers are accurate and appropriate. Unlike the European variants, prior red-teaming is not listed as a requirement. Remote hourly contract, $16–22/hour, weekly pay.
$16 – $22 / HourWorldwideRed-team AI models in Dutch and English: jailbreaks, prompt injection, bias exploitation and multi-turn manipulation, written up as reproducible attack cases and labelled data. Native Dutch plus prior adversarial or security experience. Remote hourly contract at $48–62/hour, paid weekly.
$48 – $62 / HourWorldwideEngineers with propulsion, initiation or effects test experience red-team frontier AI models: write benign, dual-use and adversarial prompts, judge the replies against a policy standard, and write reference answers with the reasoning. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideUmbrella listing for Mercor's nuclear red-team panel: fuel-cycle engineers, safeguards inspectors, nuclear security, forensics and nonproliferation specialists write prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 21 hired this month.
$65 – $75 / TaskWorldwideRed-team AI chat models and agents in Thai and English, then document every jailbreak, injection or biased answer as reproducible safety data. Native Thai required, plus prior adversarial or security experience. Remote hourly contract at $24–35/hour, paid weekly; 205 hired this month.
$24 – $35 / HourWorldwideNative Thai speakers write prompts on sensitive subjects, classify conversations and flag adversarial phrasing to make AI models safer in Thai. Business English and a bachelor's (in progress is fine) required. Part-time remote at $18–22/hour; Southeast Asia preferred, not required.
$18 – $22 / HourWorldwideThe highest-paid generalist role in Mercor's bilingual AI safety series: native Norwegian speakers write sensitive-topic prompts, classify conversations and flag adversarial phrasing. Bachelor's (in progress is fine) and business English. Part-time remote at $58–62/hour; Norway preferred, not required.
$58 – $62 / HourWorldwideUkrainian-speaking PhD chemists and biologists write specialised science prompts in Ukrainian and grade AI answers for accuracy and safe handling of dual-use topics. Part-time remote at $48–52/hour, $10 above the Ukrainian generalist role. Eastern Europe preferred, not required; 11 hires this month.
$48 – $52 / HourWorldwideRadiological emergency planners, field monitoring teams and consequence modellers red-team frontier AI models: write benign, dual-use and adversarial prompts, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideJapanese-speaking PhD chemists and biologists write specialised science prompts in Japanese and grade AI answers for accuracy and dual-use safety. Part-time remote at $68–72/hour, second-highest in the series and $20 above the Japanese generalist role. East Asia preferred, not required.
$68 – $72 / HourWorldwideTest AI models for safety failures in Finnish and English: jailbreaks, prompt injection, bias and multi-turn manipulation, logged as structured red-team data. For native Finnish speakers with prior adversarial, security or abuse-analysis experience. Remote hourly contract, $48–62/hour.
$48 – $62 / HourWorldwideSwedish and English red-teaming of AI models: jailbreaks, injected instructions, bias and multi-turn manipulation, written up as reproducible attack cases. The busiest listing in this family, with 243 hires this month. Remote hourly contract at $48–62/hour, paid weekly.
$48 – $62 / HourWorldwideTest AI models in Gujarati and English for jailbreaks, bias and harmful answers, and judge whether their Gujarati is accurate and appropriate. Evaluation judgment is the core requirement. Open to the Gujarati diaspora, with no residence rule published. Remote hourly contract at $16–22/hour; 56 hired this month.
$16 – $22 / HourWorldwideProbe AI chat models and agents for safety failures in Danish and English, then turn each failure into labelled, reproducible red-team data. For native Danish speakers with a security, adversarial ML or abuse-analysis background. Remote hourly contract, $48–62/hour, weekly pay.
$48 – $62 / HourWorldwideLicensed blasters and blasting engineers red-team frontier AI models: write benign, dual-use and adversarial prompts from quarry, mine and demolition practice, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideWrite sensitive-topic prompts in Ukrainian and classify prompts and conversations for an AI safety project, flagging adversarial phrasing as you go. Native Ukrainian, business English and a bachelor's (finished or in progress). Part-time remote at $38–42/hour; Eastern Europe preferred, not required.
$38 – $42 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.