AI Safety Experts: English & Thai
- Pay
- $24 – $35 / Hour
- Open to
- Worldwide
- Apply
We earn a commission if you sign up through the links on this page. It costs you nothing and does not affect which jobs we list. How this works.
- Skills
- red teaming
- jailbreak testing
- prompt injection
- cultural sensitivity review
- thai
- data annotation
What you'll do
Thai is a hard case for AI safety systems. It is written without spaces between words, uses its own script, carries a layered register system and deep cultural sensitivities, and is thinly represented next to English in model training. That makes it fertile ground for red-teaming, and Mercor's project wants native Thai speakers to probe it.
The brief:
- Attack conversational models and agents with jailbreaks, prompt injections, misuse cases, bias exploitation and gradual multi-turn manipulation
- Turn outcomes into labelled data: annotate failures, classify vulnerabilities and flag systemic risks
- Follow the project's taxonomies, benchmarks and playbooks so results are consistent
- Write reports, datasets and attack cases that the client can reproduce
A native Thai reviewer can recognise when an answer is offensive in a Thai social or political context, when a model mishandles honorifics or register in a way that changes meaning, and when switching script or mixing in English loosens a safeguard.
All tasks are text. Expect content on sensitive subjects (bias, misinformation, harmful behaviour); topics are announced first, and higher-sensitivity work is optional with guidelines and wellness resources.
Who fits
- Native Thai and native-level English, both required
- Some prior red-teaming: AI adversarial testing, cybersecurity, or socio-technical probing
- Disciplined, framework-driven testing habits
- Clear written explanations for technical and non-technical readers
- Comfortable across changing projects and customers
The ad also welcomes adversarial ML, pentesting and reverse engineering, harassment or disinformation analysis, and creative backgrounds in psychology, acting or writing. No degree or experience years are stated.
What it pays
$24–35 per hour, Mercor's published figure. That is a clear step below the $48–62 paid for the Nordic and Dutch versions of this same listing, and above the $17–25 offered for Indonesian, Vietnamese and Malay. The template and the experience bar are the same across all of them, so the difference reflects Mercor's market rate by language rather than harder work.
Weekly payment via Stripe or Wise is stated. H-1B and STEM OPT candidates cannot be supported.
Worth knowing
Good:
- High activity: 205 hires on this listing this month
- Competitive for Thai-language remote work paid in US dollars
- No residence requirement published, so Thai speakers abroad can apply
- Optional participation in the most sensitive projects
Less good:
- Same red-teaming bar as the top-band variants for roughly half the rate
- Disturbing material is part of the job, with safeguards
- No weekly hours stated; projects can be extended, shortened or ended early
- Independent contractor, with tax and benefits on you
- Clients and target models are unnamed
About this listing
A Mercor remote hourly contract, read from the Mercor site on 24 September 2026. Pay, language requirement, payment terms and visa exclusion are the ad's own. Mercor states the work involves no confidential information from any employer, client or institution. See Mercor.
More roles at Mercor
See all 304Similar roles at other platforms
Guides about Mercor
See all 52- How much does Mercor pay? Rates from 304 live listingsPay breakdown
- The Mercor AI interview: what happens, what it checks, and what comes afterInterview prep
- Mercor vs Alignerr: which to apply to, and how they differPlatform comparison
- Mercor, micro1, Outlier, Alignerr and Handshake AI compared: which to apply to firstPlatform comparison
Browse similar roles
Not the right fit?
See every open role, or get new ones on Telegram or Discord as they are added.