PhD chemists and biologists who write fluent Finnish: author specialised science prompts and judge how AI models answer them, including where a question strays into dual-use territory. Part-time remote at $61–65/hour, $13 above the Finnish generalist role. Finland or Western Europe preferred, not required.
Remote AI training and data labeling jobs
Filter jobs
Location: Worldwide
- Worldwide58 jobs, applied. Activate to remove
- United States5 jobs
- United Kingdom1 job
- Canadano roles alongside your other filters
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
- English56 jobs
- Germanno roles alongside your other filters
- Spanishno roles alongside your other filters
- Frenchno roles alongside your other filters
- Japanese1 job
- Portuguese1 job
Field: Science & Math
- Languages & Linguistics111 jobs
- Audio & Voice98 jobs
- Engineering91 jobs
- Business & Finance74 jobs
- Software & IT70 jobs
- Health & Medicine62 jobs
- Law, Policy & Security51 jobs
- General & Data Collection48 jobs
- Science & Math58 jobs, applied. Activate to remove
- AI Safety & Evaluation59 jobs
- Video, Image & Design34 jobs
- Data, AI & ML27 jobs
- Writing & Education11 jobs
- Other fields6 jobs
Level
- Entryno roles alongside your other filters
- Junior1 job
- Medium16 jobs
- Senior41 jobs
Newest
58 open roles matching these filters · page 2 of 3
- $61 – $65 / HourWorldwide
Thai-speaking PhD chemists and biologists write specialised science prompts in Thai and grade how AI models answer them, with a focus on dual-use safety. Part-time remote at $24–28/hour, $6 above the Thai generalist role. Southeast Asia preferred, not required; 10 hires this month.
$24 – $28 / HourWorldwideResearch-level benchmark work on bacterial population growth: two-state growth-rate switching with gamma-distributed waiting times, the Euler-Lotka equation, renewal theory and first-passage times. Solver, Auditor or Adjudicator roles. Ten openings, contractor, $80–160/hour, remote.
$80 – $160 / HourWorldwideCurate and annotate drug-discovery datasets, judge AI answers on medicinal chemistry and molecular analysis, and write feedback that improves the model. Contractor, remote, 30 openings, $90–120/hour. An advanced degree and cheminformatics or omics experience are the preferred background.
$90 – $120 / HourWorldwideFor Danish-speaking PhD scientists in chemistry or biology: write specialised prompts in Danish, grade AI answers for accuracy and safe handling, and classify conversations against guidelines. Part-time remote at $61–65/hour. Denmark or Western Europe preferred, not required; PhD candidates eligible.
$61 – $65 / HourWorldwideFormulation and synthesis chemists from pyrotechnics or propellant work red-team frontier AI models: write benign, dual-use and adversarial prompts, grade the model replies against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideThe best-paid role in Mercor's bilingual AI safety series: Norwegian-speaking PhD chemists and biologists write specialised science prompts and grade how AI models handle dual-use questions. Part-time remote at $77–81/hour. Norway preferred, not required; PhD candidates eligible.
$77 – $81 / HourWorldwideCroatian-speaking PhD chemists and biologists write specialised science prompts in Croatian and evaluate how AI models answer, including how they handle dual-use questions. Part-time remote at $48–52/hour, $10 above the Croatian generalist role. Eastern Europe preferred, not required; 9 hires this month.
$48 – $52 / HourWorldwideAuthor or verify expert multiple-choice biology questions (one correct answer, nine subtle distractors, chain-of-thought solution, references) for an AI benchmark in pharma manufacturing, synthetic biology, drug discovery and agricultural, environmental and food biology. PhD or candidate preferred. $60–75/hour, 10+ hours a week.
$60 – $75 / HourWorldwideQuantum and computational chemistry PhDs write original, runnable research problems for a scientific-computing AI benchmark, with grading criteria, calibrated until frontier models fail them more often than they pass. 6 weeks at 20+ hours a week, Git and Docker workflow. Flat $70/hour; 182 hired this month.
$70 / HourWorldwideHands-on preclinical scientists from companies that develop their own drugs annotate R&D data and model outputs, and advise an AI lab building foundation models for drug discovery. Any modality, 5+ years industry R&D, about 10 hours a week. $60–100/hour.
$60 – $100 / HourWorldwidePreclinical scientists with direct antibody-drug conjugate or bispecific antibody experience annotate R&D data and advise an AI lab building drug discovery foundation models. In-house asset developers only, 5+ years industry R&D, about 10 hours a week. $60–100/hour.
$60 – $100 / HourWorldwideRed-team frontier AI models from the chemical defence side: write benign, dual-use and adversarial prompts drawn from countermeasures, protection and detection work, judge how models respond against a policy standard, and write the reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwideKorean-speaking PhD chemists and biologists write specialised science prompts in Korean and grade AI answers for accuracy and dual-use safety. Part-time remote at $63–67/hour, $15 above the Korean generalist role and third-highest in the series. East Asia preferred, not required.
$63 – $67 / HourWorldwideHindi-speaking PhD chemists and biologists write specialised science prompts in Hindi and grade AI answers for accuracy and safe handling of dual-use topics. Part-time remote at $23–27/hour; India or South Asia preferred, not required. 12 hires this month, among the busiest in the series.
$23 – $27 / HourWorldwidePortuguese-speaking PhD chemists and biologists write specialised science prompts in Portuguese and grade how AI models handle accuracy and dual-use safety. Part-time remote at $50–54/hour; Portugal or Western Europe preferred, not required. PhD candidates eligible; 9 hires this month.
$50 – $54 / HourWorldwideArabic-speaking PhD chemists and biologists write specialised science prompts in Arabic and grade AI answers for accuracy and dual-use safety. Part-time remote at $38–42/hour; Saudi Arabia or MENA preferred, not required. 13 hires this month, the most active listing in the series.
$38 – $42 / HourWorldwideUmbrella listing for Mercor's energetic materials red-team panel: chemists and engineers or operators write benign, dual-use and adversarial prompts, judge frontier AI replies against a policy standard, and write reference answers. Remote contract at $65–75 per task; 16 hired this month.
$65 – $75 / TaskWorldwideWrite or verify 10-option multiple-choice benchmark questions in applied maths (signal processing, actuarial science, optimization, climate modeling and more), with chain-of-thought solutions and references. For maths PhDs and doctoral candidates. Remote, 10+ hours a week, $61–77/hour.
$61 – $77 / HourWorldwideAuthor executable scientific-computing problems in ecology, biochemistry and genetics for Sci Code, a new AI benchmark: source a paper, dataset or repo, write the prompt and grading criteria, and keep it only if frontier models mostly fail. PhD plus Python or R, Git and Docker. 6 weeks, 20+ hours a week, $70/hour.
$70 / HourWorldwideSolve, audit or adjudicate research-level benchmark problems in quantum optics: cascaded optical parametric amplifiers, SU(1,1) interferometers, loss and two-mode squeezing, using Bogoliubov and covariance-matrix methods. Ten openings, contractor, $80–160/hour, fully remote.
$80 – $160 / HourWorldwideTurn hands-on CRISPR, cloning and genotyping know-how into expert content and assessments that train AI on molecular and cell biology. Contractor, remote, 100 openings, $70–90/hour. A bachelor's degree plus real wet-lab gene-editing experience is the bar.
$70 – $90 / HourWorldwideBuild realistic biostatistics benchmark tasks (clinical trial analysis, regulatory review, observational studies) with real datasets and 35+ item grading rubrics. Paid per accepted task, quoted at $60–100/hour. MS or PhD plus four years in pharma, CRO, hospital or academic medicine. 50 openings.
$60 – $100 / HourWorldwideRed-team frontier AI models on chemical misuse: write benign, dual-use and adversarial prompts from forensic casework, judge the model's answers against a policy standard, and write reference answers. Remote contract at $65–75 per task.
$65 – $75 / TaskWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.