Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.

Filter jobs

Location: United States
Language
  • English32 jobs
  • Germanno roles alongside your other filters
  • Spanishno roles alongside your other filters
  • Frenchno roles alongside your other filters
  • Japaneseno roles alongside your other filters
  • Portugueseno roles alongside your other filters
Show all 49 language options
Field
Level: Medium

32 open roles matching these filters · page 1 of 2

  • US-based pharmacists and pharmacy technicians listen to short audio clips of spoken medication names and score whether each is pronounced accurately and clearly, using a rubric. Remote contract at $70/hour for an AI healthcare project, with a calibration exercise first.

    $70 / HourOpen to United States
  • Review datasets and task outputs against rubrics and guidelines, flag errors and inconsistencies, and document your reasoning where the rules run out. US-based only. 30 openings, remote contractor, $30–60/hour, 3–5+ years in data quality, QA or analysis preferred.

    $30 – $60 / HourOpen to United States
  • Former management consultants rebuild and critique PowerPoint slides for storyline, structure and visual clarity, so AI learns to produce client-ready decks. At least two years client-facing at a consulting firm required. US, UK or Canada. 15 openings, contractor, $150–350/hour.

    $150 – $350 / HourOpen to United States, United Kingdom and 1 more country
  • Broker-dealer and RIA compliance staff review mock periodic reviews, Reg BI rollovers, marketing pieces and alternatives books, producing rubric-graded compliance deliverables for AI training. US-based, about 15 hours a week, listed at $100–130/hour. Needs 3+ years in US securities compliance.

    $100 – $130 / HourOpen to United States
  • Author AI evaluation tasks from real permit drawings, documents and site photos, with the correct correction letter, review decision or markup as the answer. For ICC-certified plans examiners and third-party plan reviewers with 3+ years. US only, $45–60/hour.

    $45 – $60 / HourOpen to United States
  • Turn everyday accounting work into AI training data: design scenarios from your own practice, review AI outputs for GAAP/IFRS accuracy and judgment, and write structured feedback. Any specialty, from audit and tax to bookkeeping and forensics. US only, 3+ years and a CPA, CA, ACCA, CMA or EA required. $80/hour.

    $80 / HourOpen to United States
  • Audit AWS serverless and infrastructure-as-code tasks for a frontier AI lab: multi-service designs (Lambda, API Gateway, DynamoDB, EventBridge, Step Functions), CDK or CloudFormation fidelity, IAM boundaries and retry semantics. For US engineers with 3+ years building on AWS. $70–90/hour.

    $70 – $90 / HourOpen to United States
  • US civil engineers, land development engineers and licensed surveyors author AI evaluation tasks from real drawings, aerials, maps and site documents: plan review comment sets, survey exception schedules, earthwork checks. PE or PLS and 3+ years in the seat (8+ preferred). Remote hourly contract at $45–60/hour.

    $45 – $60 / HourOpen to United States
  • US-based full-stack engineers with 3+ years build working software on a leading AI lab's pre-release models, integrating APIs, tool interfaces and evaluation harnesses and reporting where the models fail. Two language ecosystems required. Full-time W-2 through Cincinnatus LLC, 40 hours a week, $50–65/hour.

    $50 – $65 / HourOpen to United States
  • Work mock AML, KYC, sanctions and fraud cases end to end and produce the analyst deliverable (alert dispositions, SAR narratives, remediation lists), graded against a rubric to build AI training data. US-based, about 15 hours a week, $75–100/hour. Needs 3+ years in financial crime.

    $75 – $100 / HourOpen to United States
  • Former DoD professionals with 3+ years: write golden-response WARs, SITREPs, AARs and executive summaries, draft grading rubrics for military writing, and evaluate AI-generated reports. US only. 20 openings, contractor, $40–80/hour.

    $40 – $80 / HourOpen to United States
  • Use your professional eye for visual presentation to make product design and taste judgments for a top AI company's research project. For designers who work in Figma, Sketch or Adobe and know slides and docs. US, UK or Canada; 15+ hours a week; $80–180/hour.

    $80 – $180 / HourOpen to United States, United Kingdom and 1 more country
  • Model an infrastructure transaction in Excel from a synthetic document pack, then peer-review two other contractors' models with a scorecard. For associates at infrastructure PE funds with ~2 years of IB first and exposure to data centres, fibre or towers. US or UK only, about 20–25 hours over 1.5–2 weeks, $80–100/hour.

    $80 – $100 / HourOpen to United States, United Kingdom
  • Run 20-minute structured interviews that vet senior engineers (GPU kernels, security research, ML compilers) for a frontier lab's AI training and evaluation panel, then tier each candidate and write a summary. For experienced recruiters and technical vetters. US, part-time, about 10 hours a week, $50–60/hour.

    $50 – $60 / HourOpen to United States
  • Full-time W-2 coordinator (via Cincinnatus LLC) placed with a leading AI lab's GenAI team, running day-to-day operations on AI training-data projects: turning leadership input into expert guidelines, onboarding experts, tracking quality and deliverables. US-based, 3+ years of coordination plus hands-on AI data experience. $45–55/hour.

    $45 – $55 / HourOpen to United States
  • Pediatric acute care RNs evaluate AI outputs built from nursing flowsheet documentation, annotate pediatric assessment data and help write clinical benchmarks. US licence outside California, inpatient bedside work within the past 10 years. $55–65/hour.

    $55 – $65 / HourOpen to United States
  • Inpatient registered nurses with broad clinical exposure support ongoing clinical AI product work: annotation, clinical review, model evaluation, user-feedback investigation and guideline development, alongside engineers and clinicians. US only, 10 hours a week minimum, $55–65/hour.

    $55 – $65 / HourOpen to United States
  • Design finance Excel tasks from your own work (three-statement models, LBO and DCF valuation, forecasting), write model solutions, and evaluate AI attempts for a leading tech company's GenAI team. US only, W-2 through Cincinnatus LLC, 40 hours a week to end of September then 20. $70–100/hour.

    $70 – $100 / HourOpen to United States
  • Audit end-to-end AI-assisted coding sessions (traces from tools like Cursor, Copilot or Claude Code) used to train and evaluate a frontier lab's models, judging correctness, workflow and reasoning with rubric-based feedback. 3+ years of software development plus hands-on agentic coding. US-only, $70–90/hour.

    $70 – $90 / HourOpen to United States
  • ML systems engineers write and evaluate training tasks for a frontier lab across GPU kernels, performance profiling, distributed debugging and LLM inference serving, plus the rubrics that grade them. 2+ years of hands-on ML infrastructure work. Canada, UK or US; 40 hours a week. $90–120/hour.

    $90 – $120 / HourOpen to Canada, United Kingdom and 1 more country
  • Audit applied machine-learning tasks used to train and evaluate a frontier AI lab's models: experiment design, model selection, evaluation methodology, leakage and metric gaming. For US practitioners with 3+ years of hands-on experimental ML in PyTorch, TensorFlow, scikit-learn or XGBoost. $70–90/hour.

    $70 – $90 / HourOpen to United States
  • Design Excel tasks from cross-industry experience (data models, dashboards, Power Query, VBA automation), write the solutions, and evaluate AI attempts for a tech company's GenAI team. Breadth across functions, teaching or a PhD help. US only, W-2 via Cincinnatus LLC, 40 then 20 hours a week. $70–100/hour.

    $70 – $100 / HourOpen to United States
  • Turn ambiguous AI program requirements into clear, contradiction-free rater guidelines and rubrics across finance, retail, insurance, legal and sports. For linguists, instructional designers and technical writers with 3+ years and GenAI/RLHF guideline experience. US, 35+ hours a week, $45–65/hour.

    $45 – $65 / HourOpen to United States

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.