Skip to content
Labeling Jobs

AI Safety Practitioner

Pay
$60 – $70 / Hour
Open to

Albania, Austria, Belgium, Bosnia & Herzegovina, Bulgaria, Croatia, Czechia, Denmark

and 32 more countries

Estonia, Finland, France, Germany, Greece, Hungary, Iceland, Ireland, Italy, Kosovo, Latvia, Liechtenstein, Lithuania, Luxembourg, Malta, Moldova, Monaco, Netherlands, North Macedonia, Norway, Poland, Portugal, Romania, San Marino, Serbia, Slovakia, Slovenia, Spain, Sweden, Switzerland, United Kingdom, United States

Apply

We earn a commission if you sign up through the links on this page. It costs you nothing and does not affect which jobs we list. How this works.

Skills
  • ai safety
  • trust and safety
  • content policy
  • rlhf
  • rubric-based evaluation
  • fact-checking

What you'll do

You judge how frontier AI models behave when the question is neither clearly fine nor clearly forbidden. The ad calls these "grey-area" topics, and names the territory: misinformation, political persuasion, self-harm, violence, cyber, biosecurity and other sensitive domains.

For each response you assess safety, factual accuracy, policy compliance and overall quality, flag unsafe outputs, hallucinations, reasoning failures and policy violations, and write structured feedback. You also apply and help refine the rubrics used for RLHF, SFT and safety benchmarking, and work alongside AI researchers and safety teams.

The core skill is consistency. Two reviewers looking at the same borderline response should reach the same verdict for the same reasons, and that is harder than spotting obvious violations. Expect to spend real time reading policy documents and calibrating against examples before your judgments count.

Be realistic about the content, too. Reviewing self-harm and violence material at volume is emotionally heavy, and the ad does not mention wellbeing support, rotation or exposure limits. Ask about that before you commit.

Who fits

Required:

  • Bachelor's or higher in journalism, communications, psychology, sociology, public policy, law, biology, chemistry, computer science or a related field
  • 5+ years in AI safety, trust and safety, journalism, public policy, scientific research, security or similar
  • Excellent written English, critical thinking and analytical reasoning
  • The ability to evaluate nuanced, policy-sensitive scenarios consistently

Preferred: prior AI safety, RLHF, SFT or evaluation work; familiarity with safety policies, content moderation or rubric development; experience with high-risk or ambiguous content.

The degree list is deliberately wide. A former fact-checker, a policy analyst and a biosecurity researcher all qualify, and a mixed pool of backgrounds is presumably the point.

Where you can be based

The ad publishes a hard location list of 40 countries: the United States, the United Kingdom, the EU member states it names, Iceland, Liechtenstein, Norway and Switzerland, the microstates Monaco and San Marino, and several non-EU countries in southeast and eastern Europe (Albania, Bosnia and Herzegovina, Kosovo, Moldova, North Macedonia, Serbia). Canada and Australia are not on it.

What it pays

$60–70 per hour, Mercor's published figure. Payment is weekly via Stripe or Wise. H-1B and STEM OPT candidates cannot be supported. Weekly hours are not stated.

Worth knowing

Good:

  • Broad eligibility: many degree fields and a long country list
  • Substantive work shaping how models handle contested topics
  • Direct collaboration with safety researchers, per the ad

Less good:

  • 5+ years is a firm bar that rules out early-career moderators
  • Distressing content is part of the job and no support measures are described
  • Narrow band ($60–70) with no stated route to higher rates
  • No hours, duration or end client disclosed
  • Independent contractor; projects can be extended, shortened or ended early

About this listing

Posted by Mercor as a remote hourly contract, confirmed open on 24 September 2026. The pay, country list and qualifications are the ad's own. See Mercor.

More roles at Mercor

See all 304

Similar roles at other platforms

Guides about Mercor

See all 52

Browse similar roles

Not the right fit?

See every open role, or get new ones on Telegram or Discord as they are added.