The top tier of Mercor's embedded legal expert role: a full-time W-2 job (via Cincinnatus LLC) with a leading AI lab in the Bay Area, reviewing legal model outputs, writing golden solutions and building benchmarks. For partners and general counsel. Hybrid, 6-month initial term, $100–150/hour.
Remote AI training and data labeling jobs
Filter jobs
Location
- Worldwide192 jobs
- United States80 jobs
- United Kingdom11 jobs
- Canada6 jobs
- India1 job
- Mexicono roles alongside your other filters
Language: English
Field
- Languages & Linguistics50 jobs
- Audio & Voice35 jobs
- Engineering40 jobs
- Business & Finance50 jobs
- Software & IT31 jobs
- Health & Medicine26 jobs
- Law, Policy & Security17 jobs
- General & Data Collection38 jobs
- Science & Math41 jobs
- AI Safety & Evaluation61 jobs
- Video, Image & Design9 jobs
- Data, AI & ML14 jobs
- Writing & Education1 job
- Other fieldsno roles alongside your other filters
Newest
281 open roles matching these filters · page 4 of 12
- $100 – $150 / HourHybridOpen to United States
Ukrainian-speaking PhD chemists and biologists write specialised science prompts in Ukrainian and grade AI answers for accuracy and safe handling of dual-use topics. Part-time remote at $48–52/hour, $10 above the Ukrainian generalist role. Eastern Europe preferred, not required; 11 hires this month.
$48 – $52 / HourWorldwideRadiological emergency planners, field monitoring teams and consequence modellers red-team frontier AI models: write benign, dual-use and adversarial prompts, judge model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideA hybrid, Bay Area-based W-2 role embedded with a leading AI lab: senior software engineers vet model outputs, write instruction specs and golden solutions, and build engineering benchmarks. 4+ years, senior-level progression and a CS or engineering degree. 40 hours a week for an initial 6 months, $65–105/hour.
$65 – $105 / HourHybridOpen to United StatesAudit Brazilian Portuguese speech data for Amazon's Sonic project: check transcripts against audio with fixed error codes and pass/fail verdicts, and fix word-level timestamp alignment. Native Portuguese as spoken in Brazil is a hard requirement; rationales are written in English. Remote hourly contract at $25/hour.
$25 / HourWorldwideQuality-check AI-narrated audiobooks in German (Germany): pinpoint mispronounced, missing or extra words, wrongly read numbers and abbreviations, unnatural prosody and audio artefacts, and judge the overall listen. For native German audiobook listeners. Remote hourly contract, about 20 hours a week, $15–20/hour.
$15 – $20 / HourWorldwideJapanese-speaking PhD chemists and biologists write specialised science prompts in Japanese and grade AI answers for accuracy and dual-use safety. Part-time remote at $68–72/hour, second-highest in the series and $20 above the Japanese generalist role. East Asia preferred, not required.
$68 – $72 / HourWorldwideRemote hourly contract for FPGA, digital design and embedded hardware engineers who already run Intel/Altera Quartus II on their own Windows PC. $55–65/hour, paid weekly via Stripe or Wise, $10 more than the sister Vivado listing. Needs a display above 2.5 megapixels.
$55 – $65 / HourWorldwideNative or near-native English editors, proofreaders and linguists evaluate LLM-generated English, rewrite it to a high stylistic standard, annotate grammatical and stylistic features, and give feedback that guides model training at a leading AI lab. Remote, flexible hours, $50/hour.
$50 / HourWorldwideA paid expert conversation for engineers who build and run LLM agents in production: a short AI screening interview (no coding), then, if selected, a 30-minute live call on agent reliability, evaluation and internal adoption, paid $100–500 depending on depth of experience.
$100 – $500 / TaskWorldwidePaid pilot for a biotech and pharma research team: create and critique rubrics that assess commercial drugs and development programs, judge investment-style theses, and assess 5–10 companies end to end. For specialist-fund biotech analysts with 5+ years. About 10–20 hours over 1–2 weeks, $120–200/hour.
$120 – $200 / HourWorldwideRed-team frontier AI models on chemical safety: write benign, dual-use and adversarial prompts from exposure and process-hazard work, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideFull-time Bay Area hybrid role with an AI lab's research team: review model reasoning on life sciences tasks, write golden solutions and instruction specs, and design benchmarks. For life sciences PhDs with 4+ years of substantive research experience. W-2 via Cincinnatus, $65–105/hour.
$65 – $105 / HourHybridOpen to United StatesUmbrella listing for Mercor's chemistry red-team panel: synthetic, analytical, forensic, defence and process safety chemists write benign, dual-use and adversarial prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 5 hired this month.
$65 – $75 / TaskWorldwideRed-team frontier AI models on chemical misuse: write benign, dual-use and adversarial prompts from trace analysis and method validation, judge responses against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideRemote hourly contract for FPGA and RTL engineers who already have AMD/Xilinx Vivado running on their own Windows PC. $45–55/hour, paid weekly via Stripe or Wise. You supply the tool, the hardware and a screen above 2.5 megapixels; the tasks themselves are not described.
$45 – $55 / HourWorldwideWork mock AML, KYC, sanctions and fraud cases end to end and produce the analyst deliverable (alert dispositions, SAR narratives, remediation lists), graded against a rubric to build AI training data. US-based, about 15 hours a week, $75–100/hour. Needs 3+ years in financial crime.
$75 – $100 / HourOpen to United StatesRemote hourly contract for music producers, beatmakers and audio engineers who already produce in FL Studio (FruityLoops) on their own Windows PC. $30–40/hour, paid weekly via Stripe or Wise. You need your own copy and a display above 2.5 megapixels; tasks are not described.
$30 – $40 / HourWorldwideJudge AI-generated Thai song lyrics for a leading AI lab: compare them with published songs for similarity, rate quality, creativity, prompt adherence and originality, and check that slang and word choice sound natural. For Thai songwriters, lyricists, performers or music journalists. Remote, flexible, up to 6 months, $18/hour.
$18 / HourWorldwideAuthor AI evaluation tasks from real fire and life safety review work: egress plan checks, sprinkler and alarm review, hazmat control areas, firestop photo verification. For US fire marshals, fire protection engineers and NICET III+ designer-reviewers with 3+ years in the seat. Remote hourly contract at $45–60/hour.
$45 – $60 / HourOpen to United StatesEvaluate AI-narrated audiobooks in Dutch (Netherlands): mark each mispronounced, skipped or added word, misread number, awkward intonation and audio glitch, then rate the overall listen. For native Dutch speakers who are regular audiobook listeners. Remote hourly contract, about 20 hours a week, $15–20/hour.
$15 – $20 / HourWorldwideListen to AI-narrated audiobooks in Brazilian Portuguese and tag where the synthetic narrator slips: mispronunciations, skipped or extra words, wrong readings of numbers and abbreviations, flat or odd intonation. For native Brazilian audiobook listeners. Remote hourly contract, around 20 hours a week, $15–20/hour.
$15 – $20 / HourWorldwideAuthor AI evaluation tasks in budgeting, forecasting, variance analysis and capital allocation: realistic FP&A scenarios, reference models and decks, and rubrics that reward real planning judgment over template work. For FP&A directors and CFOs with 5+ years. $80–90/hour, remote contract.
$80 – $90 / HourWorldwideTest AI models for safety failures in Finnish and English: jailbreaks, prompt injection, bias and multi-turn manipulation, logged as structured red-team data. For native Finnish speakers with prior adversarial, security or abuse-analysis experience. Remote hourly contract, $48–62/hour.
$48 – $62 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.