Audit Kubernetes tasks used to train and evaluate a frontier AI lab's models: cluster-operations scenarios, manifest and Helm correctness, and failure-mode troubleshooting (CrashLoopBackOff, OOMKilled, eviction). For US engineers with 3+ years of production Kubernetes and Go, Python or TypeScript. $70–90/hour.
Remote AI training and data labeling jobs
Filter jobs
Location: United States
- Worldwide173 jobs
- United States32 jobs, applied. Activate to remove
- United Kingdom6 jobs
- Canada5 jobs
- India1 job
- Mexicono roles alongside your other filters
Language
- English32 jobs
- Germanno roles alongside your other filters
- Spanishno roles alongside your other filters
- Frenchno roles alongside your other filters
- Japaneseno roles alongside your other filters
- Portugueseno roles alongside your other filters
Field
- Languages & Linguistics1 job
- Audio & Voice1 job
- Engineering3 jobs
- Business & Finance12 jobs
- Software & IT7 jobs
- Health & Medicine4 jobs
- Law, Policy & Security3 jobs
- General & Data Collection4 jobs
- Science & Mathno roles alongside your other filters
- AI Safety & Evaluationno roles alongside your other filters
- Video, Image & Design1 job
- Data, AI & ML3 jobs
- Writing & Education1 job
- Other fieldsno roles alongside your other filters
Newest
32 open roles matching these filters · page 2 of 2
- $70 – $90 / HourOpen to United States
Evaluate vulnerability-reproduction and remediation tasks for a frontier AI lab: faithful CVE reproductions in Docker labs, sound fixes, and two-part verification (functionality plus vulnerability tests). For US AppSec engineers, pentesters and vulnerability researchers with 3+ years. $70–90/hour.
$70 – $90 / HourOpen to United StatesAuthor AI evaluation tasks from real drawing sets, documents and site photos, with the correct RFI response, coordination comments or markup as the answer. For licensed architects, project architects and job captains with 3+ years. US only, $45–60/hour.
$45 – $60 / HourOpen to United StatesComplete self-contained fund-ops exercises from mock ledgers, statements and notices: reconciliations, NAV variance attribution, trade-break investigations, corporate actions and payment exceptions, each graded against a rubric. 3+ years in fund admin or investment ops. US only, about 15 hours a week, $75–110/hour.
$75 – $110 / HourOpen to United StatesBuild a real estate deal model in Excel from a synthetic document pack, with live formulas and sourced assumptions, then score two other contractors' models against fixed criteria. For associates at real-estate-focused mid-market PE funds with ~2 years of IB first. US only, about 20–25 hours over 1.5–2 weeks, $80–100/hour.
$80 – $100 / HourOpen to United StatesTurn everyday insurance judgment into AI training data: design realistic underwriting, claims, actuarial and compliance scenarios, review AI outputs and write feedback. Open to every insurance specialty, 3+ years, US-based. $75/hour.
$75 / HourOpen to United StatesScore and compare AI-generated writing, then write the rationale that explains each judgement. micro1 calls the exercise a Write-like-Human Eval: you are the reader whose standards the model gets trained against. Ten openings, contractor, $90–140/hour, open to the US, Canada and the UK.
$90 – $140 / HourOpen to United States, Canada and 1 more countryReview everyday professional work (documents, slides, spreadsheets) and judge it for accuracy, clarity, and completeness, writing out your reasoning. No AI background needed, but a bachelor’s degree and US or Canada residency are hard requirements.
$50 – $70 / HourOpen to United States, Canada
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.