Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.

Filter jobs

Location: United States
Language
Field
Level

81 open roles matching these filters · page 1 of 4

  • US-based pharmacists and pharmacy technicians listen to short audio clips of spoken medication names and score whether each is pronounced accurately and clearly, using a rubric. Remote contract at $70/hour for an AI healthcare project, with a calibration exercise first.

    $70 / HourOpen to United States
  • Broker-dealer and RIA compliance staff review mock periodic reviews, Reg BI rollovers, marketing pieces and alternatives books, producing rubric-graded compliance deliverables for AI training. US-based, about 15 hours a week, listed at $100–130/hour. Needs 3+ years in US securities compliance.

    $100 – $130 / HourOpen to United States
  • Author AI evaluation tasks from real permit drawings, documents and site photos, with the correct correction letter, review decision or markup as the answer. For ICC-certified plans examiners and third-party plan reviewers with 3+ years. US only, $45–60/hour.

    $45 – $60 / HourOpen to United States
  • Turn everyday accounting work into AI training data: design scenarios from your own practice, review AI outputs for GAAP/IFRS accuracy and judgment, and write structured feedback. Any specialty, from audit and tax to bookkeeping and forensics. US only, 3+ years and a CPA, CA, ACCA, CMA or EA required. $80/hour.

    $80 / HourOpen to United States
  • Evaluate frontier AI responses on grey-area and policy-sensitive topics (misinformation, political persuasion, self-harm, violence, cyber, biosecurity), apply safety rubrics and write structured feedback. 5+ years in trust and safety, journalism, policy, research or security. US, UK and most of Europe. $60–70/hour.

    $60 – $70 / HourOpen to Albania, Austria and 38 more countries
  • Audit AWS serverless and infrastructure-as-code tasks for a frontier AI lab: multi-service designs (Lambda, API Gateway, DynamoDB, EventBridge, Step Functions), CDK or CloudFormation fidelity, IAM boundaries and retry semantics. For US engineers with 3+ years building on AWS. $70–90/hour.

    $70 – $90 / HourOpen to United States
  • US civil engineers, land development engineers and licensed surveyors author AI evaluation tasks from real drawings, aerials, maps and site documents: plan review comment sets, survey exception schedules, earthwork checks. PE or PLS and 3+ years in the seat (8+ preferred). Remote hourly contract at $45–60/hour.

    $45 – $60 / HourOpen to United States
  • A paid ~20-minute online survey for current or former Devoted Health Medicare Advantage members aged 64+ in the US, mixing multiple-choice and short voice answers to inform healthcare AI tools. Listed at $120/hour, paid once after verification.

    $120 / HourOpen to United States
  • Record one session of about four hours so a Mercor client can clone your voice for its internal customer-experience AI agent. Native American English, US-based, home studio setup required; no acting experience needed. $50–150/hour, the widest band in Mercor's CX voice cloning family. You are licensing your voice, so read the usage terms first.

    $50 – $150 / HourOpen to United States
  • Judge how well AI assistants handle real personal tasks (meal planning, health habits, careers, learning) for an AI research client. For heavy everyday users of ChatGPT, Claude, Gemini and similar tools. US only, 20–40 hours a week, $50–200/hour.

    $50 – $200 / HourOpen to United States
  • Generate, structure and evaluate expert ALD and thin-film data for a frontier AI lab building semiconductor and physical-science models: solve hard problems, rate model reasoning, and turn recipes into model-ready data. US-based, 10–40 hours a week, $84/hour.

    $84 / HourOpen to United States
  • Full-time W-2 role through Cincinnatus LLC, embedded with a leading AI lab: senior board-certified physicians write golden solutions, instruction specs and clinical benchmarks for frontier models. Hybrid in the Bay Area, 40 hours a week for an initial 6 months, 4+ years post-residency and a US licence. $70–110/hour.

    $70 – $110 / HourHybridOpen to United States
  • Join a leading AI lab's research team as its accounting and audit specialist: review model outputs for misapplied standards, write instruction specs and golden solutions, and design benchmarks. Needs an active CPA, CA, CIA, CFE or CMA and 4+ years; bookkeeping-only roles do not count. Full-time W-2, hybrid Bay Area, $60–100/hour.

    $60 – $100 / HourHybridOpen to United States
  • US-based full-stack engineers with 3+ years build working software on a leading AI lab's pre-release models, integrating APIs, tool interfaces and evaluation harnesses and reporting where the models fail. Two language ecosystems required. Full-time W-2 through Cincinnatus LLC, 40 hours a week, $50–65/hour.

    $50 – $65 / HourOpen to United States
  • Build enterprise legal evaluation tasks for AI: realistic Fortune 500 scenarios, model-grade reference work and criterion-referenced rubrics. For in-house counsel or AmLaw 100 lawyers with an active bar admission, US-based, 20+ hours a week. $110–150/hour.

    $110 – $150 / HourOpen to United States
  • Residency-trained physicians in any specialty write grading criteria, evaluate multi-turn clinical dialogues and annotate clinical reasoning for AI systems. Shared pool with several workstreams. US only, 3+ years post-residency, 20 hours a week minimum, $150/hour.

    $150 / HourOpen to United States
  • Practising US hospitalists and inpatient internists review H&Ps, progress notes and discharge summaries, and judge AI-written inpatient documentation against what a hospitalist would chart. Needs C1 or better in one of 25 listed languages. 2+ years post-residency, 10 hours a week, a flat $170/hour.

    $170 / HourOpen to United States
  • Build and evaluate training data for a frontier lab's materials science models: DFT, AIMD, classical MD, surface and adsorption modeling, reaction energetics. For US-based computational PhDs fluent in VASP, Quantum ESPRESSO, CP2K, LAMMPS or ASE. Long-term, 10–40 hours a week, $84/hour.

    $84 / HourOpen to United States
  • Full-time W-2 role (via Cincinnatus LLC) for counsel and senior associates with 8 to 15 years' practice, embedded with a leading AI lab in the Bay Area to review legal model outputs, write golden solutions and build benchmarks. Hybrid, 6-month initial term, $85–120/hour. US bar admission required.

    $85 – $120 / HourHybridOpen to United States
  • Write point-in-time forecasts on specific swing-state Senate, governor and statewide races, and grade AI political analyses against your own. For state pollsters, campaign analysts and political scientists with live-race experience. US or Canada residents, $150–250/hour.

    $150 – $250 / HourOpen to Canada, United States
  • The top tier of Mercor's embedded legal expert role: a full-time W-2 job (via Cincinnatus LLC) with a leading AI lab in the Bay Area, reviewing legal model outputs, writing golden solutions and building benchmarks. For partners and general counsel. Hybrid, 6-month initial term, $100–150/hour.

    $100 – $150 / HourHybridOpen to United States
  • A hybrid, Bay Area-based W-2 role embedded with a leading AI lab: senior software engineers vet model outputs, write instruction specs and golden solutions, and build engineering benchmarks. 4+ years, senior-level progression and a CS or engineering degree. 40 hours a week for an initial 6 months, $65–105/hour.

    $65 – $105 / HourHybridOpen to United States
  • Full-time Bay Area hybrid role with an AI lab's research team: review model reasoning on life sciences tasks, write golden solutions and instruction specs, and design benchmarks. For life sciences PhDs with 4+ years of substantive research experience. W-2 via Cincinnatus, $65–105/hour.

    $65 – $105 / HourHybridOpen to United States

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.