Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
  • Work through real property management tasks in Buildium on screen recording (tenants, leases, rent, maintenance), then write prompts and rubrics from each workflow to train AI agents. Needs your own Buildium access and prior human-data work. 40 openings, $28–92/hour.

    $28 – $92 / HourWorldwide
  • Build features, fix hard bugs and refactor legacy code inside existing open source projects, with clear write-ups of each solution, as training input for AI coding models. Expert level in two languages plus open source and competitive programming experience. 37 openings, paid per task on a $50–100/hour band.

    $50 – $100 / HourWorldwide
  • Former HUMINT collectors, investigative interviewers and negotiators: run or assess structured interview sessions, read verbal and nonverbal cues, rate reliability and annotate sessions against a rubric to train AI. At least about 15 hours a week, US-based preferred. 10 openings, contractor, $25–65/hour.

    $25 – $65 / HourWorldwide
  • Judge whether Hebrew speech, human or AI-generated, sounds genuinely native, and explain each rating in written English. Native Hebrew and B2 English; phonetics or voice work helps but no AI experience is needed. 100 openings, contractor, $30–65/hour.

    $30 – $65 / HourWorldwide
  • Record and edit professional Thai speech audio for an AI training dataset, with accurate tones and minimal background noise, delivered to strict naming and format specs. Three years of recording and editing experience preferred; gear needed at interview. 25 openings, contractor, $10–30/hour.

    $10 – $30 / HourWorldwide
  • Build, tune and evaluate ML models in Python with MongoDB-backed data pipelines for a customer AI training project, and document every experiment. Remote contractor, $80–140/hour, 35 openings. scikit-learn, TensorFlow or PyTorch plus hands-on MongoDB is the core stack.

    $80 – $140 / HourWorldwide
  • Read clinical images of skin lesions, describe them in precise terms, judge how well AI assessments hold up, and write the criteria that define a high-quality dermatological read. Non-clinical, no patient care. Active US licence and five years post-residency. A flat $270/hour.

    $270 / HourOpen to United States
  • The UK posting of Mercor's molecular biology project: design primers, plasmids, gRNAs, mRNA constructs and repair templates as ground truth for a frontier model, and write the rubrics that judge them. PhD strongly preferred, first-author record expected, 20 hours a week. $70–105/hour.

    $70 – $105 / HourOpen to United Kingdom
  • Design the primers, plasmids, gRNAs, mRNA constructs and repair templates that become ground truth for a frontier model, and write the rubrics that judge sequence design quality. PhD strongly preferred, first-author record expected, 20 hours a week minimum. US only, $70–105/hour.

    $70 – $105 / HourOpen to United States
  • Contribute your own eligible images, describe them from personal knowledge, and judge how well an AI photo product understands them and responds to your comments and corrections. Entry to the project is by an eligibility survey. $17/hour, remote independent contractor work through Meridial.

    $17 / HourWorldwide
  • Technical experts audit the tasks used to train and evaluate AI systems, checking each is accurate, realistic, solvable, reproducible and properly tested. You are matched to one of eight specialties, from GPU kernels and Kubernetes to CVE security and SWE-Bench, and take an assessment in it. $60/hour, remote contractor.

    $60 / HourWorldwide
  • Review work product against rubrics and quality standards, document what is wrong with it, and write the feedback that fixes it. It sits a layer above annotation: you check other people's output. Remote hourly contract at $20–50/hour, paid weekly, no degree stated.

    $20 – $50 / HourWorldwide
  • Build, break and verify real machine learning engineering tasks: model components, reproducible training and inference workflows, memory and throughput optimisation, numerical debugging. Output-based pay, 100 openings, global remote, roughly 15 hours a week at $100–150/hour.

    $100 – $150 / HourWorldwide
  • Grade AI assistant outputs against detailed rubrics at volume, find the reasoning gaps, tool-use failures and logic errors, and write the feedback that fixes them. No degree requirement stated. Ten openings, contractor, $30–90/hour, six eligible countries.

    $30 – $90 / HourOpen to United States, Canada and 4 more countries
  • Score AI-generated specs, release notes, user-facing copy and stakeholder updates, then write the rationale behind each score. The ad is unusually blunt that this is evaluation work only: no roadmap, no delivery, no sprint facilitation. Ten openings, contractor, remote, $90–140/hour.

    $90 – $140 / HourWorldwide
  • Score and compare AI-generated writing, then write the rationale that explains each judgement. micro1 calls the exercise a Write-like-Human Eval: you are the reader whose standards the model gets trained against. Ten openings, contractor, $90–140/hour, open to the US, Canada and the UK.

    $90 – $140 / HourOpen to United States, Canada and 1 more country
  • Run simulated contract negotiations and redlining exercises, review how an AI handles the same scenarios, and write the grading rubrics that measure it. Three years of in-house technology-transactions experience is the stated bar. Task-based pay at roughly 3.5 hours per task.

    $90 – $130 / HourWorldwide
  • Assess adolescent eating disorder cases in writing, applying DSM-5, EDE-Q and SCOFF, judging medical stability against MEED, and saying whether FBT or CBT-E fits the clinical stage. Wants a licensed clinician with a recent adolescent caseload, and accepts paediatric medicine alongside psychiatry and psychology. Remote contractor work.

    $45 – $95 / HourWorldwide
  • Solve hard problems in your own scientific field, build the reference solution in Python, and judge whether an AI's answer is actually correct, catching the wrong assumption or missing constraint rather than just the wrong number. Seven fields qualify, and no specific degree or number of years is required. Global, roughly 15 hours a week.

    $50 – $100 / HourWorldwide
  • Refine CRM data models, arbitrate conflicting operational policies and document why you ruled the way you did, and design revenue workflows that hold up under vertical regulation. Four years with your hands actually in the system is the bar; the ad states outright that management-only exposure does not count. Remote contractor work.

    $50 – $100 / HourWorldwide
  • Redline commercial agreements (MSAs, NDAs, DPAs) against established playbooks, decide the company position and the acceptable fallbacks, and write the reasoning behind every markup. Asks for a JD, active US bar admission and four years of transactional practice. Remote contractor work on a project described only as confidential.

    $90 – $110 / HourWorldwide
  • Advise on ab initio configuration interaction calculations of polarizability, critique atomic-structure work on Yb-171 and Yb-174, and explain sigma versus pi transitions, largely through spoken AI interviews, which on this listing are the work rather than the screening. Worldwide, 5 to 10 hours a week, and the ad asks only for what you can legally share.

    $100 – $200 / HourWorldwide
  • Run simulated contract negotiations and redlining exercises, judge how an AI handles contract scenarios, and build the grading frameworks that measure it. Three years in-house on technology transactions (MSAs, NDAs, DPAs) is the stated bar, and the ad spells out what task-based pay means: roughly 3.5 hours per task.

    $100 – $130 / HourWorldwide

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.