Review datasets and task outputs against rubrics and guidelines, flag errors and inconsistencies, and document your reasoning where the rules run out. US-based only. 30 openings, remote contractor, $30–60/hour, 3–5+ years in data quality, QA or analysis preferred.
Remote AI training and data labeling jobs
Every role here has been checked against the platform that posted it. Pay is shown as reported, and marked when it is an estimate rather than a firm rate.
Filter jobs
Location: United States
- Worldwide9 jobs
- United States3 jobs, applied. Activate to remove
- United Kingdom1 job
- Canada1 job
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
- English3 jobs
- Germanno roles alongside your other filters
- Spanishno roles alongside your other filters
- Frenchno roles alongside your other filters
- Japaneseno roles alongside your other filters
- Portugueseno roles alongside your other filters
Field: Data, AI & ML
- Languages & Linguistics1 job
- Audio & Voice1 job
- Engineering3 jobs
- Business & Finance12 jobs
- Software & IT7 jobs
- Health & Medicine4 jobs
- Law, Policy & Security3 jobs
- General & Data Collection4 jobs
- Science & Mathno roles alongside your other filters
- AI Safety & Evaluationno roles alongside your other filters
- Video, Image & Design1 job
- Data, AI & ML3 jobs, applied. Activate to remove
- Writing & Education1 job
- Other fieldsno roles alongside your other filters
Level: Medium
- Entryno roles alongside your other filters
- Juniorno roles alongside your other filters
- Medium3 jobs, applied. Activate to remove
- Senior4 jobs
Newest
3 open roles matching these filters
- $30 – $60 / HourOpen to United States
ML systems engineers write and evaluate training tasks for a frontier lab across GPU kernels, performance profiling, distributed debugging and LLM inference serving, plus the rubrics that grade them. 2+ years of hands-on ML infrastructure work. Canada, UK or US; 40 hours a week. $90–120/hour.
$90 – $120 / HourOpen to Canada, United Kingdom and 1 more countryAudit applied machine-learning tasks used to train and evaluate a frontier AI lab's models: experiment design, model selection, evaluation methodology, leakage and metric gaming. For US practitioners with 3+ years of hands-on experimental ML in PyTorch, TensorFlow, scikit-learn or XGBoost. $70–90/hour.
$70 – $90 / HourOpen to United States
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.