AI Safety Experts: English & Marathi
- Pay
- $16 – $22 / Hour
- Open to
- Worldwide
- Apply
We earn a commission if you sign up through the links on this page. It costs you nothing and does not affect which jobs we list. How this works.
- Skills
- red teaming
- ai response evaluation
- content review
- language quality review
- marathi
- data annotation
What you'll do
Marathi shares the Devanagari script with Hindi, and that is exactly where AI models tend to go wrong. With far more Hindi than Marathi in their training data, models often drift into Hindi vocabulary, grammar or cultural assumptions while claiming to answer in Marathi. A native Marathi reader spots that immediately; an automated check usually does not. The same blind spots are where safety behaviour weakens.
Mercor's ad lists red-team tasks:
- Probe chat models and agents using jailbreaks, prompt injections, misuse cases, bias exploitation and multi-turn manipulation
- Annotate failures, classify vulnerabilities and flag systemic risks
- Follow taxonomies, benchmarks and playbooks
- Deliver reproducible reports, datasets and attack cases
The "who you are" section for Marathi is evaluation-led: strong judgment about whether an AI response is accurate, complete and appropriate, a habit of catching subtle mistakes, consistent use of guidelines, and clear explanations. Prior red-teaming, which the European versions of this listing expect, is not in the core profile here.
All work is text. Content covers sensitive areas such as bias, misinformation and harmful behaviour; topics are disclosed up front, and higher-sensitivity projects are optional with guidelines and wellness resources.
Who fits
- Native fluency in Marathi and English
- A good ear for when an answer is subtly wrong, incomplete or inappropriate, and the words to explain why
- Attention to detail and consistency with quality standards
- Clear written communication
- Comfort switching between task types and clients
Pluses: adversarial ML, cybersecurity, harassment and disinformation analysis, psychology, acting or writing. No degree or minimum years are stated.
What it pays
$16–22 per hour, Mercor's published band, the same as the other Indian-language variants and Urdu. The ad states weekly payment via Stripe or Wise. H-1B and STEM OPT candidates are not supported.
Worth knowing
Good:
- Evaluation-focused entry bar, so careful native speakers from editing, teaching or translation backgrounds can compete
- Weekly US-dollar payouts
- No residence requirement published
- Sensitive work is opt-in, with topics announced first
Less good:
- No hires count on the listing when we read it, so demand is unclear
- Lowest pay band in this family, well below the $48–62 European rate
- Red-team content can be distressing
- No weekly hours; projects can be extended, shortened or ended early
- Contractor terms; clients and models are not named
About this listing
Posted by Mercor as a remote hourly contract, read from the Mercor site on 24 September 2026. Pay, requirements, payment terms and the visa exclusion are the ad's own. Mercor states the work involves no confidential information from any employer, client or institution. See Mercor.
More roles at Mercor
See all 304Guides about Mercor
See all 52- How much does Mercor pay? Rates from 304 live listingsPay breakdown
- The Mercor AI interview: what happens, what it checks, and what comes afterInterview prep
- Mercor vs Alignerr: which to apply to, and how they differPlatform comparison
- Mercor, micro1, Outlier, Alignerr and Handshake AI compared: which to apply to firstPlatform comparison
Browse similar roles
Not the right fit?
See every open role, or get new ones on Telegram or Discord as they are added.