Test AI chat models in Malayalam and English for safety failures and judge whether their answers are accurate, complete and appropriate. Native Malayalam plus careful evaluation skills; prior red-teaming is a plus rather than the core requirement. Remote hourly contract at $16–22/hour, paid weekly.
Remote AI training and data labeling jobs
Filter jobs
Location
Language
Field
- Languages & Linguistics58 jobs
- Audio & Voice43 jobs
- Engineering40 jobs
- Business & Finance52 jobs
- Software & IT32 jobs
- Health & Medicine28 jobs
- Law, Policy & Security17 jobs
- General & Data Collection42 jobs
- Science & Math41 jobs
- AI Safety & Evaluation61 jobs
- Video, Image & Design9 jobs
- Data, AI & ML15 jobs
- Writing & Education1 job
- Other fieldsno roles alongside your other filters
Newest
305 open roles matching these filters · page 9 of 13
- $16 – $22 / HourWorldwide
Full-time W-2 coordinator (via Cincinnatus LLC) placed with a leading AI lab's GenAI team, running day-to-day operations on AI training-data projects: turning leadership input into expert guidelines, onboarding experts, tracking quality and deliverables. US-based, 3+ years of coordination plus hands-on AI data experience. $45–55/hour.
$45 – $55 / HourOpen to United StatesPreclinical scientists with direct antibody-drug conjugate or bispecific antibody experience annotate R&D data and advise an AI lab building drug discovery foundation models. In-house asset developers only, 5+ years industry R&D, about 10 hours a week. $60–100/hour.
$60 – $100 / HourWorldwideNuclear medicine physicists, radiopharmacy staff and cyclotron RSOs red-team frontier AI models: write benign, dual-use and adversarial prompts from medical isotope work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwidePredict when a named sell-side analyst will publish after a catalyst and what the note will say, reasoning only from a fixed evidence cutoff, then grade AI analyses of the same call. For former lead sell-side analysts or very senior associates in the exact sector, typically 8+ years. $150–250/hour, remote.
$150 – $250 / HourWorldwideHelp run LLM training projects on browsing capabilities: track project and annotator performance in Google Sheets, review annotator quality and handle contributor communications. Ops or project management background preferred, but ownership matters more. Remote hourly contract at $50–60/hour, no location limit published.
$50 – $60 / HourWorldwideNative Telugu speakers in India map the layout of real Telugu PDF pages (headings, tables, figures, reading order) and transcribe every text region exactly in Telugu script, including handwriting, to train document AI. Remote hourly contract at $12.68/hour, with peer review of every task.
$12.68 / HourOpen to IndiaPediatric acute care RNs evaluate AI outputs built from nursing flowsheet documentation, annotate pediatric assessment data and help write clinical benchmarks. US licence outside California, inpatient bedside work within the past 10 years. $55–65/hour.
$55 – $65 / HourOpen to United StatesBelgium-based PhD chemists and biologists who write Belgian Dutch: author specialised science prompts and grade how AI models handle accuracy and dual-use safety. Belgium residence required. Part-time remote at $61–65/hour, $13 above the Belgian Dutch generalist role; 9 hires this month.
$61 – $65 / HourOpen to BelgiumNative Danish speakers write sensitive-topic prompts, classify prompts and conversations, and flag adversarial phrasing to make AI models safer in Danish. Business English and a bachelor's (in progress counts). Part-time remote at $48–52/hour; Denmark or Western Europe preferred, not required.
$48 – $52 / HourWorldwideAustralia-based native English speakers with an authentic Australian accent record one session of about four hours, which a Mercor client uses to clone the voice for its internal customer-experience AI agent. $50–100/hour in USD, no acting experience required. You are licensing your voice, so check the terms before recording.
$50 – $100 / HourOpen to AustraliaRead a specific non-G7 central bank's statements, minutes and speeches in the source language, score them dovish to hawkish from a fixed evidence cutoff, and grade AI macro analyses. For former central bank economists or senior local rates and FX strategists, typically 8+ years. $150–250/hour, remote.
$150 – $250 / HourWorldwideSpain-based native speakers of Castilian (Peninsular) Spanish record one session of about four hours, which a Mercor client uses to clone the voice for its internal customer-experience AI agent. $50–100/hour in USD, no acting experience required. You are licensing your voice, so ask for the usage terms first.
$50 – $100 / HourOpen to SpainAudit French speech data for Amazon's Sonic project: check annotators' transcripts against audio with fixed error codes and pass/fail calls, and correct word-level timestamp alignment. Native French as spoken in France is a hard requirement; rationales are in English. Joint top rate of the Sonic set at $41.50/hour, remote contract.
$41.50 / HourWorldwideRadioactive source security and vulnerability assessment specialists red-team frontier AI models: write benign, dual-use and adversarial prompts from Category 1 and 2 source security work, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideUS-based cloud and DevOps engineers design and grade AI training tasks on Kubernetes failure diagnosis, AWS service integration, Terraform or CDK design and CI/CD, and write the rubrics behind them. 4+ years at a top-tier organisation. Full-time W-2 through Cincinnatus LLC at a leading AI lab, $75–110/hour.
$75 – $110 / HourOpen to United StatesInpatient registered nurses with broad clinical exposure support ongoing clinical AI product work: annotation, clinical review, model evaluation, user-feedback investigation and guideline development, alongside engineers and clinicians. US only, 10 hours a week minimum, $55–65/hour.
$55 – $65 / HourOpen to United StatesDesign finance Excel tasks from your own work (three-statement models, LBO and DCF valuation, forecasting), write model solutions, and evaluate AI attempts for a leading tech company's GenAI team. US only, W-2 through Cincinnatus LLC, 40 hours a week to end of September then 20. $70–100/hour.
$70 – $100 / HourOpen to United StatesSouth Africa-based native English speakers with an authentic South African accent record one session of about four hours, which a Mercor client uses to clone the voice for its internal customer-experience AI agent. $50–100/hour in USD, a strong rate locally. No acting experience required. You are licensing your voice.
$50 – $100 / HourOpen to South AfricaFemale native Peruvian Spanish speakers living in Peru record one session of about four hours, which a Mercor client uses to clone the voice for its internal customer-experience AI agent. $50–100/hour in USD, no acting experience required; one hire already this month. You are licensing your voice.
$50 – $100 / HourOpen to PeruNative Gujarati speakers in India annotate the structure of real Gujarati PDF pages and transcribe every text region exactly in Gujarati script, handwriting included, to build training data for document AI. Remote hourly contract at $12.68/hour; every task is peer-reviewed by a second Gujarati expert.
$12.68 / HourOpen to IndiaAudit end-to-end AI-assisted coding sessions (traces from tools like Cursor, Copilot or Claude Code) used to train and evaluate a frontier lab's models, judging correctness, workflow and reasoning with rubric-based feedback. 3+ years of software development plus hands-on agentic coding. US-only, $70–90/hour.
$70 – $90 / HourOpen to United StatesAudit Italian speech data for Amazon's Sonic collections: judge annotators' transcripts against audio with fixed error codes and pass/fail verdicts, and correct word-level timestamp alignment. Native Italian as spoken in Italy is a hard requirement; rationales are written in English. Remote hourly contract at $39.50/hour.
$39.50 / HourWorldwideFull-time Bay Area hybrid role embedded with an AI lab: review model reasoning on materials problems, write golden solutions and specs, and build benchmarks. For materials PhDs (or master's with exceptional industrial depth) with 4+ years of R&D. W-2 via Cincinnatus, $70–110/hour.
$70 – $110 / HourHybridOpen to United States
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.