Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.

Filter jobs

Location
  • Worldwide21 jobs
  • United Statesno roles alongside your other filters
  • United Kingdomno roles alongside your other filters
  • Canadano roles alongside your other filters
  • Indiano roles alongside your other filters
  • Mexicono roles alongside your other filters
Show all 72 location options
Language
Show all 49 language options
Field: AI Safety & Evaluation
Level: Medium

22 open roles matching these filters

  • Red-team AI models in European and other non-Brazilian Portuguese plus English: jailbreaks, prompt injection, bias and manipulation, logged as reproducible safety data. Brazilian Portuguese is explicitly excluded. Remote hourly contract at $29–45/hour, paid weekly.

    $29 – $45 / HourWorldwide
  • Red-team conversational AI models in Norwegian and English: attempt jailbreaks, prompt injections and multi-turn manipulation, then document what broke. Prior red-teaming, security or adversarial testing experience expected. Remote hourly contract at $48–62/hour, paid weekly.

    $48 – $62 / HourWorldwide
  • Red-team AI models in Dutch and English: jailbreaks, prompt injection, bias exploitation and multi-turn manipulation, written up as reproducible attack cases and labelled data. Native Dutch plus prior adversarial or security experience. Remote hourly contract at $48–62/hour, paid weekly.

    $48 – $62 / HourWorldwide
  • Red-team AI chat models and agents in Thai and English, then document every jailbreak, injection or biased answer as reproducible safety data. Native Thai required, plus prior adversarial or security experience. Remote hourly contract at $24–35/hour, paid weekly; 205 hired this month.

    $24 – $35 / HourWorldwide
  • Ukrainian-speaking PhD chemists and biologists write specialised science prompts in Ukrainian and grade AI answers for accuracy and safe handling of dual-use topics. Part-time remote at $48–52/hour, $10 above the Ukrainian generalist role. Eastern Europe preferred, not required; 11 hires this month.

    $48 – $52 / HourWorldwide
  • Japanese-speaking PhD chemists and biologists write specialised science prompts in Japanese and grade AI answers for accuracy and dual-use safety. Part-time remote at $68–72/hour, second-highest in the series and $20 above the Japanese generalist role. East Asia preferred, not required.

    $68 – $72 / HourWorldwide
  • Test AI models for safety failures in Finnish and English: jailbreaks, prompt injection, bias and multi-turn manipulation, logged as structured red-team data. For native Finnish speakers with prior adversarial, security or abuse-analysis experience. Remote hourly contract, $48–62/hour.

    $48 – $62 / HourWorldwide
  • Swedish and English red-teaming of AI models: jailbreaks, injected instructions, bias and multi-turn manipulation, written up as reproducible attack cases. The busiest listing in this family, with 243 hires this month. Remote hourly contract at $48–62/hour, paid weekly.

    $48 – $62 / HourWorldwide
  • Probe AI chat models and agents for safety failures in Danish and English, then turn each failure into labelled, reproducible red-team data. For native Danish speakers with a security, adversarial ML or abuse-analysis background. Remote hourly contract, $48–62/hour, weekly pay.

    $48 – $62 / HourWorldwide
  • PhD chemists and biologists who write fluent Finnish: author specialised science prompts and judge how AI models answer them, including where a question strays into dual-use territory. Part-time remote at $61–65/hour, $13 above the Finnish generalist role. Finland or Western Europe preferred, not required.

    $61 – $65 / HourWorldwide
  • Thai-speaking PhD chemists and biologists write specialised science prompts in Thai and grade how AI models answer them, with a focus on dual-use safety. Part-time remote at $24–28/hour, $6 above the Thai generalist role. Southeast Asia preferred, not required; 10 hires this month.

    $24 – $28 / HourWorldwide
  • For Danish-speaking PhD scientists in chemistry or biology: write specialised prompts in Danish, grade AI answers for accuracy and safe handling, and classify conversations against guidelines. Part-time remote at $61–65/hour. Denmark or Western Europe preferred, not required; PhD candidates eligible.

    $61 – $65 / HourWorldwide
  • The best-paid role in Mercor's bilingual AI safety series: Norwegian-speaking PhD chemists and biologists write specialised science prompts and grade how AI models handle dual-use questions. Part-time remote at $77–81/hour. Norway preferred, not required; PhD candidates eligible.

    $77 – $81 / HourWorldwide
  • Adversarial testing of AI chat models in Bahasa Indonesia and English: jailbreaks, prompt injection, bias and multi-turn manipulation, recorded as structured red-team data. Native Indonesian plus prior red-teaming or security experience. Remote hourly contract, $17–25/hour, weekly pay.

    $17 – $25 / HourWorldwide
  • Probe AI chat models and agents for safety failures in Vietnamese and English (jailbreaks, prompt injection, bias, multi-turn manipulation) and document each as reproducible data. Native Vietnamese plus prior red-teaming or security experience. Remote hourly contract, $17–25/hour.

    $17 – $25 / HourWorldwide
  • Croatian-speaking PhD chemists and biologists write specialised science prompts in Croatian and evaluate how AI models answer, including how they handle dual-use questions. Part-time remote at $48–52/hour, $10 above the Croatian generalist role. Eastern Europe preferred, not required; 9 hires this month.

    $48 – $52 / HourWorldwide
  • Red-team AI models in Malay and English: jailbreaks, prompt injection, bias and multi-turn manipulation, captured as labelled, reproducible safety data. Native Malay required, along with prior adversarial, security or abuse-analysis experience. Remote hourly contract at $17–25/hour, paid weekly.

    $17 – $25 / HourWorldwide
  • Belgium-based PhD chemists and biologists who write Belgian Dutch: author specialised science prompts and grade how AI models handle accuracy and dual-use safety. Belgium residence required. Part-time remote at $61–65/hour, $13 above the Belgian Dutch generalist role; 9 hires this month.

    $61 – $65 / HourOpen to Belgium
  • Korean-speaking PhD chemists and biologists write specialised science prompts in Korean and grade AI answers for accuracy and dual-use safety. Part-time remote at $63–67/hour, $15 above the Korean generalist role and third-highest in the series. East Asia preferred, not required.

    $63 – $67 / HourWorldwide
  • Hindi-speaking PhD chemists and biologists write specialised science prompts in Hindi and grade AI answers for accuracy and safe handling of dual-use topics. Part-time remote at $23–27/hour; India or South Asia preferred, not required. 12 hires this month, among the busiest in the series.

    $23 – $27 / HourWorldwide
  • Portuguese-speaking PhD chemists and biologists write specialised science prompts in Portuguese and grade how AI models handle accuracy and dual-use safety. Part-time remote at $50–54/hour; Portugal or Western Europe preferred, not required. PhD candidates eligible; 9 hires this month.

    $50 – $54 / HourWorldwide
  • Arabic-speaking PhD chemists and biologists write specialised science prompts in Arabic and grade AI answers for accuracy and dual-use safety. Part-time remote at $38–42/hour; Saudi Arabia or MENA preferred, not required. 13 hires this month, the most active listing in the series.

    $38 – $42 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.