Portuguese-speaking PhD chemists and biologists write specialised science prompts in Portuguese and grade how AI models handle accuracy and dual-use safety. Part-time remote at $50–54/hour; Portugal or Western Europe preferred, not required. PhD candidates eligible; 9 hires this month.
Remote AI training and data labeling jobs
Filter jobs
Location
- Worldwide60 jobs
- United States28 jobs
- United Kingdom4 jobs
- Canada3 jobs
- India1 job
- Mexicono roles alongside your other filters
Language
Field
- Languages & Linguistics15 jobs
- Audio & Voice5 jobs
- Engineering11 jobs
- Business & Finance21 jobs
- Software & IT17 jobs
- Health & Medicine6 jobs
- Law, Policy & Security4 jobs
- General & Data Collection23 jobs
- Science & Math13 jobs
- AI Safety & Evaluation22 jobs
- Video, Image & Design1 job
- Data, AI & ML3 jobs
- Writing & Education1 job
- Other fieldsno roles alongside your other filters
Newest
91 open roles matching these filters · page 4 of 4
- $50 – $54 / HourWorldwide
India-based full-stack engineers build real applications on top of a leading AI lab's pre-release models, wiring in tool interfaces, evaluation harnesses and telemetry, and writing up the model failures they hit. 3+ years at a top-tier organisation and two programming languages. Full-time, 40 hours a week, $25–30/hour.
$25 – $30 / HourOpen to IndiaEvaluate generative music AI for a leading AI lab: compare AI-made songs head to head on musicality, prompt adherence, vocals and mix, label genre and structure, and check lyrics and vocals. For Thai-speaking producers or engineers with 2+ years' experience. Remote, flexible hours, up to 6 months, $18/hour.
$18 / HourWorldwideRate AI-generated music for a leading AI lab: head-to-head song comparisons on musicality, prompt adherence, vocals and mix, genre and structure labelling, and lyric and vocal checks. For Russian-speaking producers and mix engineers with 2+ years' experience. Remote, flexible hours, up to 6 months, $35–49/hour.
$35 – $49 / HourWorldwideTurn ambiguous AI program requirements into clear, contradiction-free rater guidelines and rubrics across finance, retail, insurance, legal and sports. For linguists, instructional designers and technical writers with 3+ years and GenAI/RLHF guideline experience. US, 35+ hours a week, $45–65/hour.
$45 – $65 / HourOpen to United StatesAudit Kubernetes tasks used to train and evaluate a frontier AI lab's models: cluster-operations scenarios, manifest and Helm correctness, and failure-mode troubleshooting (CrashLoopBackOff, OOMKilled, eviction). For US engineers with 3+ years of production Kubernetes and Go, Python or TypeScript. $70–90/hour.
$70 – $90 / HourOpen to United StatesExperienced Spanish tutors lead recorded one-on-one lessons with native English-speaking learners so a language-learning platform can capture authentic tutoring conversations. 3+ years' teaching and a degree required. About 7 hours a week for roughly three weeks, $60–65/hour, remote.
$60 – $65 / HourWorldwideArabic-speaking PhD chemists and biologists write specialised science prompts in Arabic and grade AI answers for accuracy and dual-use safety. Part-time remote at $38–42/hour; Saudi Arabia or MENA preferred, not required. 13 hires this month, the most active listing in the series.
$38 – $42 / HourWorldwideEvaluate vulnerability-reproduction and remediation tasks for a frontier AI lab: faithful CVE reproductions in Docker labs, sound fixes, and two-part verification (functionality plus vulnerability tests). For US AppSec engineers, pentesters and vulnerability researchers with 3+ years. $70–90/hour.
$70 – $90 / HourOpen to United StatesAuthor AI evaluation tasks from real drawing sets, documents and site photos, with the correct RFI response, coordination comments or markup as the answer. For licensed architects, project architects and job captains with 3+ years. US only, $45–60/hour.
$45 – $60 / HourOpen to United StatesNative or near-native German editors, critics and linguists evaluate LLM-generated German text, rewrite it to literary and cultural standards, annotate linguistic features and give feedback to researchers at a leading AI lab. Remote hourly contract at $50/hour.
$50 / HourWorldwideRemote hourly contract for Python developers who already run PyCharm on their own Mac. $55–65/hour, paid weekly via Stripe or Wise. You supply the licence, the Mac and a display above 2.5 megapixels. Tasks are not described in the ad.
$55 – $65 / HourWorldwideNative or near-native Spanish translators, editors and annotators evaluate LLM-generated Spanish, correct or rewrite it, annotate grammatical and semantic features, and flag model errors for researchers at a leading AI lab. Remote hourly contract at $50/hour.
$50 / HourWorldwideBuild hard codebase-exploration tasks for an RL environment that trains AI agents: rewrite engineering questions about production Go repos (etcd, Traefik, Helm, gRPC-Go, Temporal and more) so frontier agents fail them, tune rubrics and foils, and pass a validation loop. $130 per approved task, remote.
$130 / TaskWorldwideComplete self-contained fund-ops exercises from mock ledgers, statements and notices: reconciliations, NAV variance attribution, trade-break investigations, corporate actions and payment exceptions, each graded against a rubric. 3+ years in fund admin or investment ops. US only, about 15 hours a week, $75–110/hour.
$75 – $110 / HourOpen to United StatesRemote hourly contract for developers on any stack who already use Visual Studio Code on their own Mac. $55–65/hour, paid weekly via Stripe or Wise. You need your own Mac and a display above 2.5 megapixels. Tasks are not described in the ad.
$55 – $65 / HourWorldwideBuild a real estate deal model in Excel from a synthetic document pack, with live formulas and sourced assumptions, then score two other contractors' models against fixed criteria. For associates at real-estate-focused mid-market PE funds with ~2 years of IB first. US only, about 20–25 hours over 1.5–2 weeks, $80–100/hour.
$80 – $100 / HourOpen to United StatesTurn everyday insurance judgment into AI training data: design realistic underwriting, claims, actuarial and compliance scenarios, review AI outputs and write feedback. Open to every insurance specialty, 3+ years, US-based. $75/hour.
$75 / HourOpen to United StatesReview everyday professional work (documents, slides, spreadsheets) and judge it for accuracy, clarity, and completeness, writing out your reasoning. No AI background needed, but a bachelor’s degree and US or Canada residency are hard requirements.
$50 – $70 / HourOpen to United States, Canada
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.