Skip to content
Labeling Jobs

Data, AI & ML AI training jobs

Open AI training, data labeling, and annotation roles in Data, AI & ML, with the pay each platform reports and the countries it accepts.

Open roles

  • Annotate video captured by robotic systems (objects, movements, events), catch edge cases, and help the AI team refine annotation protocols and capture strategy. For people with hands-on computer vision and video annotation experience. 300 openings, remote contractor, $50–90/hour.

    $50 – $90 / HourWorldwide
  • Review datasets and task outputs against rubrics and guidelines, flag errors and inconsistencies, and document your reasoning where the rules run out. US-based only. 30 openings, remote contractor, $30–60/hour, 3–5+ years in data quality, QA or analysis preferred.

    $30 – $60 / HourOpen to United States
  • Data scientists and analysts analyse datasets, build analytics reports and rebuild data-driven PowerPoint decks so AI learns to turn numbers into clear business insight. US, UK or Canada. 7 openings, contractor, $150–350/hour.

    $150 – $350 / HourOpen to United States, United Kingdom and 1 more country
  • Remote hourly contract for biostatisticians, epidemiologists and applied economists who already run Stata SE on their own Windows PC. $45–55/hour, paid weekly via Stripe or Wise. Your own Stata SE licence and a display above 2.5 megapixels are required. Tasks are not described.

    $45 – $55 / HourWorldwide
  • Contributors aged 50 or over submit 30 unedited selfies and 10 short videos of themselves from the past five years for AI training. A one-off submission paid as a single payment; micro1's card shows $139–140/hour. 5,000 openings, remote.

    $139 – $140 / HourWorldwide
  • A full-time research post at micro1: design the benchmarks, rubrics and evaluation protocols that measure how AI agents handle enterprise finance work, and run original research on financial reasoning. Advanced degree in finance or economics required, PhD strongly preferred. Base salary $200,000–250,000 plus equity.

    $200000 – $250000 / YearWorldwide
  • Review large business datasets for accuracy and completeness under strict PII and privacy rules, flag quality problems and write findings that feed AI training. Remote contractor, $31–60/hour, 50 openings.

    $31 – $60 / HourWorldwide
  • A talent pool, not an open project: Mercor is collecting data scientists for future work evaluating how well AI does real data science, from writing grading criteria for analyses, models and A/B write-ups to scoring and justifying. 1+ year of experience at a top tech, AI or quant firm. $100–150/hour when work exists.

    $100 – $150 / HourWorldwide
  • Pick information-rich charts from your field and write unambiguous, multi-step quantitative questions about them, with worked answers, to train AI on chart reasoning. Suits medicine, finance, engineering, earth science, data science and operations professionals with two years of chart-heavy work. 50 openings, $25–50/hour.

    $25 – $50 / HourWorldwide
  • Run realistic analytics workflows through AI assistants connected to Snowflake, then check every figure they produce with your own SQL. You also maintain the seeded test data, warehouse roles and connector setups. Remote contractor, $50–60/hour, 4 openings, 3+ years of SQL and Snowflake preferred.

    $50 – $60 / HourWorldwide
  • Write point-in-time forecasts on specific swing-state Senate, governor and statewide races, and grade AI political analyses against your own. For state pollsters, campaign analysts and political scientists with live-race experience. US or Canada residents, $150–250/hour.

    $150 – $250 / HourOpen to Canada, United States
  • A paid expert conversation for engineers who build and run LLM agents in production: a short AI screening interview (no coding), then, if selected, a 30-minute live call on agent reliability, evaluation and internal adoption, paid $100–500 depending on depth of experience.

    $100 – $500 / TaskWorldwide
  • A full-time research post at micro1: design and own the benchmarks, clinical quality rubrics and validation protocols that measure AI medical reasoning and decision support, and research where models fail. Advanced healthcare degree (MD, DO, PhD, MPH, PharmD) and research experience required. Base salary $200,000–250,000 plus equity.

    $200000 – $250000 / YearWorldwide
  • Siblings or cousins apply together, each submitting 30 unedited selfies and 10 videos of themselves from the past five years for facial recognition training. Both must complete separate applications before onboarding. Card shows $139–140/hour; 3,997 openings.

    $139 – $140 / HourWorldwide
  • Data scientists from top tech, finance or research firms evaluate AI answers on statistics, ML and experimentation, and write expert prompts and reference solutions for a leading AI lab. Contractor, remote, 100 openings, $245–280/hour. English-speaking-country base preferred.

    $245 – $280 / HourWorldwide
  • A full-time micro1 research role designing robotics datasets, collection methods, annotation schemas and evaluations for embodied AI (egocentric video, teleop, trajectories, simulation). Base salary $200,000–320,000 plus equity, remote, one opening. Two to five-plus years in robotics or applied ML.

    $200000 – $320000 / YearWorldwide
  • Write multi-step benchmark questions on clinical and epidemiological charts (Kaplan-Meier curves, forest plots, ROC curves, hazard-ratio plots) with unambiguous, fully worked answers for AI evaluation. MD, MPH, PhD or related degree plus two years. 50 openings, contractor, $25–50/hour.

    $25 – $50 / HourWorldwide
  • Hands-on ML researchers take on scoped, open-ended empirical problems: training image classifiers and generators from scratch, fine-tuning open-weight LLMs, adversarial robustness, compression under hard budgets, and multilingual pre-training. 3+ years of ML research (PhD counts). Remote hourly contract at $100–120/hour.

    $100 – $120 / HourWorldwide
  • Author point-in-time election and political-risk forecasts, document calls on polling and win probabilities, and grade AI analyses. For senior national forecasters, campaign analytics leads, pollsters and political-risk analysts with 8+ years. Remote, $150–250/hour.

    $150 – $250 / HourWorldwide
  • Map how an upcoming catalyst should ripple through suppliers, customers, competitors and substitutes, estimate direction and magnitude from primary filings, and grade AI analyses. For senior sector PMs and lead equity analysts with 8+ years. Remote, $150–250/hour.

    $150 – $250 / HourWorldwide
  • ML systems engineers write and evaluate training tasks for a frontier lab across GPU kernels, performance profiling, distributed debugging and LLM inference serving, plus the rubrics that grade them. 2+ years of hands-on ML infrastructure work. Canada, UK or US; 40 hours a week. $90–120/hour.

    $90 – $120 / HourOpen to Canada, United Kingdom and 1 more country
  • Audit applied machine-learning tasks used to train and evaluate a frontier AI lab's models: experiment design, model selection, evaluation methodology, leakage and metric gaming. For US practitioners with 3+ years of hands-on experimental ML in PyTorch, TensorFlow, scikit-learn or XGBoost. $70–90/hour.

    $70 – $90 / HourOpen to United States
  • Evaluate Neuron Kernel Interface (NKI) development tasks for a frontier AI lab: CUDA-to-NKI migration fidelity, Trainium performance optimisation and GPU-versus-Trainium numerical correctness. Requires 2+ years writing NKI kernels for Trainium or Inferentia2. US-only, $70–90/hour.

    $70 – $90 / HourOpen to United States
  • Evaluate GPU and accelerator kernel development tasks for a frontier AI lab: numerical correctness, benchmarking fairness, task scoping and compile or runtime validity across CUDA, Triton, NKI and Pallas. For US engineers with 3+ years of kernel work in at least two of those frameworks. $70–90/hour.

    $70 – $90 / HourOpen to United States
  • Grade AI-generated slides, spreadsheets and documents for real-world data science quality, flagging factual, visual and presentation errors in structured written feedback. Needs 5+ years at a top firm in the US, UK, Canada, Australia or New Zealand. $100–150/hour.

    $100 – $150 / HourWorldwide
  • An expert-interview listing for engineers who have shipped production search, especially agentic search in the LLM era: a 25-minute conversational interview about relevance, evaluation and real trade-offs, with a possible paid 30-minute follow-up call at $200. No coding, no take-home. Listed at $80–150 per task.

    $80 – $150 / TaskWorldwide
  • Author executable scientific-computing problems in ecology, biochemistry and genetics for Sci Code, a new AI benchmark: source a paper, dataset or repo, write the prompt and grading criteria, and keep it only if frontier models mostly fail. PhD plus Python or R, Git and Docker. 6 weeks, 20+ hours a week, $70/hour.

    $70 / HourWorldwide
  • A full-time research engineering role at micro1 building RL environments, reward functions, verifiers, synthetic data pipelines and automated evaluation systems. Base salary $200,000–300,000 plus equity and benefits, remote, one opening. Deep reinforcement learning experience required.

    $200000 – $300000 / YearWorldwide
  • Write multi-step benchmark questions about hard data-science charts (Sankey diagrams, heatmaps, calibration curves, residual plots, correlation matrices) with unambiguous, fully worked answers for AI evaluation. Bachelor's or equivalent plus two years. 50 openings, contractor, $25–50/hour.

    $25 – $50 / HourWorldwide
  • Clean messy real-world datasets, run descriptive and inferential analyses in Python or R, build clear visualisations and explain the results to non-technical readers, feeding AI model training data. MS or PhD in a quantitative field preferred. 50 openings, contractor, $60–120/hour.

    $60 – $120 / HourWorldwide
  • Run realistic developer workflows (pull requests, code review, issues, CI runs) through AI tools and grade whether the results are correct, while seeding reproducible repos and wiring up integrations. Three years of Git and CI/CD experience; MCP and connector familiarity valued. Four openings, $50–70/hour.

    $50 – $70 / HourWorldwide
  • Build, tune and evaluate ML models in Python with MongoDB-backed data pipelines for a customer AI training project, and document every experiment. Remote contractor, $80–140/hour, 35 openings. scikit-learn, TensorFlow or PyTorch plus hands-on MongoDB is the core stack.

    $80 – $140 / HourWorldwide
  • Build, break and verify real machine learning engineering tasks: model components, reproducible training and inference workflows, memory and throughput optimisation, numerical debugging. Output-based pay, 100 openings, global remote, roughly 15 hours a week at $100–150/hour.

    $100 – $150 / HourWorldwide
  • Refine CRM data models, arbitrate conflicting operational policies and document why you ruled the way you did, and design revenue workflows that hold up under vertical regulation. Four years with your hands actually in the system is the bar; the ad states outright that management-only exposure does not count. Remote contractor work.

    $50 – $100 / HourWorldwide
  • Build and document the production pipelines you'd work on during a normal week, alongside experts from two other disciplines, so the outputs can be turned into the tasks and rubrics that train frontier models. $140–200/hour for 40 hours a week from 14 September to 10 October 2026. You must be based in the UK with the right to work there.

    $140 – $200 / HourOpen to United Kingdom

Jobs in other fields

Nothing that fits today?

New roles land most days. Get them on Telegram or Discord as they are added, or read how the platforms pay before you apply.