Full-time finance specialist embedded with a leading AI lab: QA model outputs, write instruction specs and golden solutions, and build finance benchmarks across FP&A, IB, asset management, PE, risk or treasury. 5+ years at a recognised institution, VP-level progression. W-2, hybrid Bay Area, $60–100/hour.
Remote AI training and data labeling jobs
Filter jobs
Location: United States
Language
Field
- Languages & Linguistics2 jobs
- Audio & Voice4 jobs
- Engineering17 jobs
- Business & Finance30 jobs
- Software & IT16 jobs
- Health & Medicine16 jobs
- Law, Policy & Security19 jobs
- General & Data Collection14 jobs
- Science & Math5 jobs
- AI Safety & Evaluation2 jobs
- Video, Image & Design9 jobs
- Data, AI & ML7 jobs
- Writing & Education3 jobs
- Other fieldsno roles alongside your other filters
Newest
118 open roles matching these filters · page 5 of 5
- $60 – $100 / HourHybridOpen to United States
Coordinate the finance raters on a leading AI lab's training-data program: build tracking and escalation workflows, monitor throughput and quality, triage rater questions and keep finance tasks consistent. 5–10 years in finance or finance operations with team coordination experience. US only, 35+ hours a week, $40–60/hour.
$40 – $60 / HourOpen to United StatesEvaluate vulnerability-reproduction and remediation tasks for a frontier AI lab: faithful CVE reproductions in Docker labs, sound fixes, and two-part verification (functionality plus vulnerability tests). For US AppSec engineers, pentesters and vulnerability researchers with 3+ years. $70–90/hour.
$70 – $90 / HourOpen to United StatesAuthor AI evaluation tasks from real drawing sets, documents and site photos, with the correct RFI response, coordination comments or markup as the answer. For licensed architects, project architects and job captains with 3+ years. US only, $45–60/hour.
$45 – $60 / HourOpen to United StatesEvaluate AI model outputs on underwriting, claims and risk reasoning against rubrics, design hard insurance tasks with worked solutions, and refine scoring guidelines. Needs 8+ years at a top-tier insurer or broker and prior hands-on LLM rubric evaluation. US only, 35+ hours a week, $60–80/hour.
$60 – $80 / HourOpen to United StatesEvaluate AI finance outputs against rubrics, design hard finance tasks with worked solutions, and refine scoring guidelines for a leading AI lab. Needs 8+ years at a top-tier bank, asset manager or Big Four firm plus prior hands-on LLM rubric evaluation. US only, 35+ hours a week, $65–90/hour.
$65 – $90 / HourOpen to United StatesCertified pharmacy technicians answer medication questions and review AI responses for an AI lab building prior authorization workflows: dosing, interactions, indications, PA requirements. US only, 30–40 hours a week during the project. A flat $35/hour.
$35 / HourOpen to United StatesComplete self-contained fund-ops exercises from mock ledgers, statements and notices: reconciliations, NAV variance attribution, trade-break investigations, corporate actions and payment exceptions, each graded against a rubric. 3+ years in fund admin or investment ops. US only, about 15 hours a week, $75–110/hour.
$75 – $110 / HourOpen to United StatesFull-time IB and M&A specialist embedded with a leading AI lab: vet model outputs on deal work, write instruction specs and golden solutions, and build finance benchmarks. 5+ years at a recognised institution, VP-level progression, MBA or CFA preferred. W-2 via Cincinnatus, hybrid Bay Area, $100–150/hour.
$100 – $150 / HourHybridOpen to United StatesBuild a real estate deal model in Excel from a synthetic document pack, with live formulas and sourced assumptions, then score two other contractors' models against fixed criteria. For associates at real-estate-focused mid-market PE funds with ~2 years of IB first. US only, about 20–25 hours over 1.5–2 weeks, $80–100/hour.
$80 – $100 / HourOpen to United StatesFull-time PE and VC specialist embedded with a leading AI lab: QA model outputs on investment work, write instruction specs and golden solutions, and design finance benchmarks. 5+ years at a recognised institution with Principal or VP-level ownership of decisions. W-2 via Cincinnatus, hybrid Bay Area, $110–150/hour.
$110 – $150 / HourHybridOpen to United StatesTurn everyday insurance judgment into AI training data: design realistic underwriting, claims, actuarial and compliance scenarios, review AI outputs and write feedback. Open to every insurance specialty, 3+ years, US-based. $75/hour.
$75 / HourOpen to United StatesFilm your everyday household chores from a first-person, head-mounted phone to train AI, paid per hour of accepted video. Portuguese-language ad for residents of 39 listed US states with US work eligibility. 1,000 openings, $10–15/hour, $20 head-strap bonus after 20 hours.
$10 – $15 / HourOpen to United StatesFilm your own household chores from a first-person view with your smartphone on a head mount, for AI training data. $13–15 per hour of accepted video. US only, and only in the states the ad lists. Needs a recent iPhone, Pixel or Galaxy.
$13 – $15 / HourOpen to United StatesRecord your household chores first-person with a smartphone on a head strap, producing training video for AI. Open in 40 named US states, not all 50. $10–15 per hour of accepted video, plus a $20 bonus after 20 hours. 1,000 openings.
$10 – $15 / HourOpen to United StatesRead clinical images of skin lesions, describe them in precise terms, judge how well AI assessments hold up, and write the criteria that define a high-quality dermatological read. Non-clinical, no patient care. Active US licence and five years post-residency. A flat $270/hour.
$270 / HourOpen to United StatesJudge AI-generated clinical notes against what a practising outpatient physician would actually document, and help write the annotation guidelines. Any specialty, but you need C1 or better in Czech, Catalan, Danish, Dutch, Vietnamese or Finnish. US only, 10 hours a week, $170–190/hour.
$170 – $190 / HourOpen to United StatesDesign the primers, plasmids, gRNAs, mRNA constructs and repair templates that become ground truth for a frontier model, and write the rubrics that judge sequence design quality. PhD strongly preferred, first-author record expected, 20 hours a week minimum. US only, $70–105/hour.
$70 – $105 / HourOpen to United StatesGrade AI assistant outputs against detailed rubrics at volume, find the reasoning gaps, tool-use failures and logic errors, and write the feedback that fixes them. No degree requirement stated. Ten openings, contractor, $30–90/hour, six eligible countries.
$30 – $90 / HourOpen to United States, Canada and 4 more countriesScore and compare AI-generated writing, then write the rationale that explains each judgement. micro1 calls the exercise a Write-like-Human Eval: you are the reader whose standards the model gets trained against. Ten openings, contractor, $90–140/hour, open to the US, Canada and the UK.
$90 – $140 / HourOpen to United States, Canada and 1 more countryReview everyday professional work (documents, slides, spreadsheets) and judge it for accuracy, clarity, and completeness, writing out your reasoning. No AI background needed, but a bachelor’s degree and US or Canada residency are hard requirements.
$50 – $70 / HourOpen to United States, CanadaWrite difficult problems in your own field for a top AI lab's language models. Open to PhDs (or equivalent industry experience) in medicine, statistics, AI/ML, computer science, game development or mechanical and aerospace engineering, living in the US, UK, Canada or Australia. Remote hourly contract at $55–80/hour, paid weekly.
$55 – $80 / HourOpen to Australia, Canada and 2 more countries
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.