Evaluate frontier AI responses on grey-area and policy-sensitive topics (misinformation, political persuasion, self-harm, violence, cyber, biosecurity), apply safety rubrics and write structured feedback. 5+ years in trust and safety, journalism, policy, research or security. US, UK and most of Europe. $60–70/hour.
Remote AI training and data labeling jobs
Filter jobs
Location: United Kingdom
- Worldwide83 jobs
- United States45 jobs
- United Kingdom6 jobs, applied. Activate to remove
- Canada2 jobs
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
- English4 jobs
- Germanno roles alongside your other filters
- Spanishno roles alongside your other filters
- Frenchno roles alongside your other filters
- Japaneseno roles alongside your other filters
- Portugueseno roles alongside your other filters
Field
- Languages & Linguisticsno roles alongside your other filters
- Audio & Voiceno roles alongside your other filters
- Engineeringno roles alongside your other filters
- Business & Finance1 job
- Software & IT1 job
- Health & Medicineno roles alongside your other filters
- Law, Policy & Securityno roles alongside your other filters
- General & Data Collection1 job
- Science & Math1 job
- AI Safety & Evaluation2 jobs
- Video, Image & Designno roles alongside your other filters
- Data, AI & ML1 job
- Writing & Educationno roles alongside your other filters
- Other fieldsno roles alongside your other filters
Newest
6 open roles matching these filters
- $60 – $70 / HourOpen to Albania, Austria and 38 more countries
Design adversarial prompts, find jailbreaks and policy failures, and document vulnerabilities in frontier AI models across cyber, biosecurity, fraud, misinformation and political content. Hourly remote contract at $70–84/hour for residents of Europe, the UK and the US; 5+ years' relevant experience required.
$70 – $84 / HourOpen to Albania, Austria and 38 more countriesThe UK posting of Mercor's molecular biology project: design primers, plasmids, gRNAs, mRNA constructs and repair templates as ground truth for a frontier model, and write the rubrics that judge them. PhD strongly preferred, first-author record expected, 20 hours a week. $70–105/hour.
$70 – $105 / HourOpen to United KingdomProduce the decks, models and operating-model work you'd build on a normal engagement, alongside experts from two other disciplines, so the outputs can be turned into the tasks and rubrics that train frontier models. $140–200/hour for 40 hours a week from 14 September to 10 October 2026. You must be based in the UK with the right to work there.
$140 – $200 / HourOpen to United KingdomBuild and document the production pipelines you'd work on during a normal week, alongside experts from two other disciplines, so the outputs can be turned into the tasks and rubrics that train frontier models. $140–200/hour for 40 hours a week from 14 September to 10 October 2026. You must be based in the UK with the right to work there.
$140 – $200 / HourOpen to United KingdomWrite difficult problems in your own field for a top AI lab's language models. Open to PhDs (or equivalent industry experience) in medicine, statistics, AI/ML, computer science, game development or mechanical and aerospace engineering, living in the US, UK, Canada or Australia. Remote hourly contract at $55–80/hour, paid weekly.
$55 – $80 / HourOpen to Australia, Canada and 2 more countries
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.