Judge AI-generated Thai song lyrics for a leading AI lab: compare them with published songs for similarity, rate quality, creativity, prompt adherence and originality, and check that slang and word choice sound natural. For Thai songwriters, lyricists, performers or music journalists. Remote, flexible, up to 6 months, $18/hour.
Remote AI training and data labeling jobs
Filter jobs
Location
Language
Field
- Languages & Linguistics58 jobs
- Audio & Voice43 jobs
- Engineering40 jobs
- Business & Finance52 jobs
- Software & IT32 jobs
- Health & Medicine28 jobs
- Law, Policy & Security17 jobs
- General & Data Collection42 jobs
- Science & Math41 jobs
- AI Safety & Evaluation61 jobs
- Video, Image & Design9 jobs
- Data, AI & ML15 jobs
- Writing & Education1 job
- Other fieldsno roles alongside your other filters
Newest
305 open roles matching these filters · page 5 of 13
- $18 / HourWorldwide
Author AI evaluation tasks from real fire and life safety review work: egress plan checks, sprinkler and alarm review, hazmat control areas, firestop photo verification. For US fire marshals, fire protection engineers and NICET III+ designer-reviewers with 3+ years in the seat. Remote hourly contract at $45–60/hour.
$45 – $60 / HourOpen to United StatesEvaluate AI-narrated audiobooks in Dutch (Netherlands): mark each mispronounced, skipped or added word, misread number, awkward intonation and audio glitch, then rate the overall listen. For native Dutch speakers who are regular audiobook listeners. Remote hourly contract, about 20 hours a week, $15–20/hour.
$15 – $20 / HourWorldwideListen to AI-narrated audiobooks in Brazilian Portuguese and tag where the synthetic narrator slips: mispronunciations, skipped or extra words, wrong readings of numbers and abbreviations, flat or odd intonation. For native Brazilian audiobook listeners. Remote hourly contract, around 20 hours a week, $15–20/hour.
$15 – $20 / HourWorldwideAuthor AI evaluation tasks in budgeting, forecasting, variance analysis and capital allocation: realistic FP&A scenarios, reference models and decks, and rubrics that reward real planning judgment over template work. For FP&A directors and CFOs with 5+ years. $80–90/hour, remote contract.
$80 – $90 / HourWorldwideTest AI models for safety failures in Finnish and English: jailbreaks, prompt injection, bias and multi-turn manipulation, logged as structured red-team data. For native Finnish speakers with prior adversarial, security or abuse-analysis experience. Remote hourly contract, $48–62/hour.
$48 – $62 / HourWorldwideSwedish and English red-teaming of AI models: jailbreaks, injected instructions, bias and multi-turn manipulation, written up as reproducible attack cases. The busiest listing in this family, with 243 hires this month. Remote hourly contract at $48–62/hour, paid weekly.
$48 – $62 / HourWorldwideAudit repository-level software engineering benchmark tasks for a frontier AI lab: reference patches, test harnesses, Docker isolation, and signs of answer leakage or reward hacking. For US engineers with 3+ years and real open-source contributor or maintainer history. $70–90/hour.
$70 – $90 / HourOpen to United StatesRemote hourly contract for mechanical designers and drafters who already run AutoCAD Mechanical on their own Windows PC for 2D drafting and detailing. $45–55/hour, paid weekly via Stripe or Wise. Plain AutoCAD may not be enough; the ad names the Mechanical toolset.
$45 – $55 / HourWorldwideTest AI models in Gujarati and English for jailbreaks, bias and harmful answers, and judge whether their Gujarati is accurate and appropriate. Evaluation judgment is the core requirement. Open to the Gujarati diaspora, with no residence rule published. Remote hourly contract at $16–22/hour; 56 hired this month.
$16 – $22 / HourWorldwideUse your professional eye for visual presentation to make product design and taste judgments for a top AI company's research project. For designers who work in Figma, Sketch or Adobe and know slides and docs. US, UK or Canada; 15+ hours a week; $80–180/hour.
$80 – $180 / HourOpen to United States, United Kingdom and 1 more countryRemote hourly contract for graphic designers, illustrators and brand or packaging designers who already use Adobe Illustrator on their own Windows PC. $30–40/hour, paid weekly via Stripe or Wise. Needs your own Adobe subscription and a display above 2.5 megapixels.
$30 – $40 / HourWorldwidePaid pilot for board-certified, practising physicians with deep therapeutic-area expertise: interpret trial endpoints, judge whether results would change prescribing and real-world uptake, and write rubrics that evaluate AI analysis of drugs. 10–20 hours over 1–2 weeks. $150–230/hour.
$150 – $230 / HourWorldwideRemote hourly contract for engineers who already use MATLAB on their own Windows PC for simulation, signal or image processing or control work. $45–55/hour, paid weekly on Stripe or Wise. Your own licence and a high-resolution display are required; the tasks are not spelled out.
$45 – $55 / HourWorldwideShare your personal ChatGPT history with an AI research organization through Mercor, then possibly join paid follow-on project work. For native Korean speakers based in South Korea with a heavily used personal account. $50/hour.
$50 / HourOpen to South KoreaGrade AI-generated slides, spreadsheets and documents to consulting standard, flagging factual, visual and presentation errors with structured written feedback. Needs 5+ years at a top firm in the US, UK, Canada, Australia or New Zealand. $100–150/hour.
$100 – $150 / HourWorldwideRed-team frontier AI models on chemical misuse: write benign, dual-use and adversarial prompts from route design and scale-up experience, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideA paid remote research study (Project Mercator) for 100 current or former US public defenders: describe the tasks you actually did in a written survey. $100 per hour of reviewed survey time, capped at 90 minutes ($150) for the first wave. No case documents needed.
$100 / HourWorldwideListen to AI-narrated Italian audiobooks and log every slip in the synthetic voice: skipped or mispronounced words, misread numbers, odd intonation, glitches. For native Italian speakers who actually listen to audiobooks. Remote hourly contract, about 20 hours a week, $15–20/hour.
$15 – $20 / HourWorldwideRemote hourly contract for economists and econometricians who already run EViews on their own Windows PC. You bring the licence and a high-resolution screen; Mercor pays $45–55/hour, weekly via Stripe or Wise. The ad does not describe the tasks, hours or project length.
$45 – $55 / HourWorldwideProbe AI chat models and agents for safety failures in Danish and English, then turn each failure into labelled, reproducible red-team data. For native Danish speakers with a security, adversarial ML or abuse-analysis background. Remote hourly contract, $48–62/hour, weekly pay.
$48 – $62 / HourWorldwideWrite and verify expert-level physics multiple-choice questions (one right answer, nine plausible wrong ones) for an AI benchmark, across semiconductors, photonics, quantum sensing, plasma, turbulence and geophysics. PhD or doctoral candidate preferred. Remote hourly contract at $61–77/hour, 10+ hours a week.
$61 – $77 / HourWorldwideDesign Fortune 500 go-to-market scenarios, write reference deal strategies and campaigns, and author rubrics that test whether AI shows real enterprise sales and marketing judgment. For senior F500 sellers, marketers and RevOps leads. Remote hourly contract at $60–70/hour.
$60 – $70 / HourWorldwideLicensed blasters and blasting engineers red-team frontier AI models: write benign, dual-use and adversarial prompts from quarry, mine and demolition practice, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.