Skip to content
Labeling Jobs

Remote AI training and data labeling jobs

Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.

Filter jobs

Location: United States
Language
  • English7 jobs
  • Germanno roles alongside your other filters
  • Spanishno roles alongside your other filters
  • Frenchno roles alongside your other filters
  • Japaneseno roles alongside your other filters
  • Portugueseno roles alongside your other filters
Show all 49 language options
Field: Software & IT
Level: Medium

7 open roles matching these filters

  • Audit AWS serverless and infrastructure-as-code tasks for a frontier AI lab: multi-service designs (Lambda, API Gateway, DynamoDB, EventBridge, Step Functions), CDK or CloudFormation fidelity, IAM boundaries and retry semantics. For US engineers with 3+ years building on AWS. $70–90/hour.

    $70 – $90 / HourOpen to United States
  • US-based full-stack engineers with 3+ years build working software on a leading AI lab's pre-release models, integrating APIs, tool interfaces and evaluation harnesses and reporting where the models fail. Two language ecosystems required. Full-time W-2 through Cincinnatus LLC, 40 hours a week, $50–65/hour.

    $50 – $65 / HourOpen to United States
  • Run 20-minute structured interviews that vet senior engineers (GPU kernels, security research, ML compilers) for a frontier lab's AI training and evaluation panel, then tier each candidate and write a summary. For experienced recruiters and technical vetters. US, part-time, about 10 hours a week, $50–60/hour.

    $50 – $60 / HourOpen to United States
  • Audit end-to-end AI-assisted coding sessions (traces from tools like Cursor, Copilot or Claude Code) used to train and evaluate a frontier lab's models, judging correctness, workflow and reasoning with rubric-based feedback. 3+ years of software development plus hands-on agentic coding. US-only, $70–90/hour.

    $70 – $90 / HourOpen to United States
  • ML systems engineers write and evaluate training tasks for a frontier lab across GPU kernels, performance profiling, distributed debugging and LLM inference serving, plus the rubrics that grade them. 2+ years of hands-on ML infrastructure work. Canada, UK or US; 40 hours a week. $90–120/hour.

    $90 – $120 / HourOpen to Canada, United Kingdom and 1 more country
  • Audit Kubernetes tasks used to train and evaluate a frontier AI lab's models: cluster-operations scenarios, manifest and Helm correctness, and failure-mode troubleshooting (CrashLoopBackOff, OOMKilled, eviction). For US engineers with 3+ years of production Kubernetes and Go, Python or TypeScript. $70–90/hour.

    $70 – $90 / HourOpen to United States
  • Evaluate vulnerability-reproduction and remediation tasks for a frontier AI lab: faithful CVE reproductions in Docker labs, sound fixes, and two-part verification (functionality plus vulnerability tests). For US AppSec engineers, pentesters and vulnerability researchers with 3+ years. $70–90/hour.

    $70 – $90 / HourOpen to United States

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.