A full-time research post at micro1: design the benchmarks, rubrics and evaluation protocols that measure how AI agents handle enterprise finance work, and run original research on financial reasoning. Advanced degree in finance or economics required, PhD strongly preferred. Base salary $200,000–250,000 plus equity.
Remote AI training and data labeling jobs
Filter jobs
Location: Worldwide
- Worldwide13 jobs, applied. Activate to remove
- United States4 jobs
- United Kingdom2 jobs
- Canada2 jobs
- Indiano roles alongside your other filters
- Mexicono roles alongside your other filters
Language
- English13 jobs
- Germanno roles alongside your other filters
- Spanishno roles alongside your other filters
- Frenchno roles alongside your other filters
- Japaneseno roles alongside your other filters
- Portugueseno roles alongside your other filters
Field: Data, AI & ML
- Languages & Linguisticsno roles alongside your other filters
- Audio & Voiceno roles alongside your other filters
- Engineering49 jobs
- Business & Finance30 jobs
- Software & IT32 jobs
- Health & Medicine54 jobs
- Law, Policy & Security35 jobs
- General & Data Collection1 job
- Science & Math41 jobs
- AI Safety & Evaluation20 jobs
- Video, Image & Design3 jobs
- Data, AI & ML13 jobs, applied. Activate to remove
- Writing & Education2 jobs
- Other fields5 jobs
Newest
13 open roles matching these filters
- $200000 – $250000 / YearWorldwide
A paid expert conversation for engineers who build and run LLM agents in production: a short AI screening interview (no coding), then, if selected, a 30-minute live call on agent reliability, evaluation and internal adoption, paid $100–500 depending on depth of experience.
$100 – $500 / TaskWorldwideA full-time research post at micro1: design and own the benchmarks, clinical quality rubrics and validation protocols that measure AI medical reasoning and decision support, and research where models fail. Advanced healthcare degree (MD, DO, PhD, MPH, PharmD) and research experience required. Base salary $200,000–250,000 plus equity.
$200000 – $250000 / YearWorldwideHands-on ML researchers take on scoped, open-ended empirical problems: training image classifiers and generators from scratch, fine-tuning open-weight LLMs, adversarial robustness, compression under hard budgets, and multilingual pre-training. 3+ years of ML research (PhD counts). Remote hourly contract at $100–120/hour.
$100 – $120 / HourWorldwideAuthor point-in-time election and political-risk forecasts, document calls on polling and win probabilities, and grade AI analyses. For senior national forecasters, campaign analytics leads, pollsters and political-risk analysts with 8+ years. Remote, $150–250/hour.
$150 – $250 / HourWorldwideMap how an upcoming catalyst should ripple through suppliers, customers, competitors and substitutes, estimate direction and magnitude from primary filings, and grade AI analyses. For senior sector PMs and lead equity analysts with 8+ years. Remote, $150–250/hour.
$150 – $250 / HourWorldwideGrade AI-generated slides, spreadsheets and documents for real-world data science quality, flagging factual, visual and presentation errors in structured written feedback. Needs 5+ years at a top firm in the US, UK, Canada, Australia or New Zealand. $100–150/hour.
$100 – $150 / HourWorldwideAn expert-interview listing for engineers who have shipped production search, especially agentic search in the LLM era: a 25-minute conversational interview about relevance, evaluation and real trade-offs, with a possible paid 30-minute follow-up call at $200. No coding, no take-home. Listed at $80–150 per task.
$80 – $150 / TaskWorldwideAuthor executable scientific-computing problems in ecology, biochemistry and genetics for Sci Code, a new AI benchmark: source a paper, dataset or repo, write the prompt and grading criteria, and keep it only if frontier models mostly fail. PhD plus Python or R, Git and Docker. 6 weeks, 20+ hours a week, $70/hour.
$70 / HourWorldwideA full-time research engineering role at micro1 building RL environments, reward functions, verifiers, synthetic data pipelines and automated evaluation systems. Base salary $200,000–300,000 plus equity and benefits, remote, one opening. Deep reinforcement learning experience required.
$200000 – $300000 / YearWorldwideBuild, tune and evaluate ML models in Python with MongoDB-backed data pipelines for a customer AI training project, and document every experiment. Remote contractor, $80–140/hour, 35 openings. scikit-learn, TensorFlow or PyTorch plus hands-on MongoDB is the core stack.
$80 – $140 / HourWorldwideBuild, break and verify real machine learning engineering tasks: model components, reproducible training and inference workflows, memory and throughput optimisation, numerical debugging. Output-based pay, 100 openings, global remote, roughly 15 hours a week at $100–150/hour.
$100 – $150 / HourWorldwideRefine CRM data models, arbitrate conflicting operational policies and document why you ruled the way you did, and design revenue workflows that hold up under vertical regulation. Four years with your hands actually in the system is the bar; the ad states outright that management-only exposure does not count. Remote contractor work.
$50 – $100 / HourWorldwide
Nothing that fits today?
New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.