AI Safety Specialist - Evaluation Expert

Mercor
Any Location, IA
Job Description
Role Overview

Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality. Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains. Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.

What You Will Do

Evaluate AI-generated responses, identify unsafe outputs, hallucinations, reasoning failures, and policy violations, and provide structured feedback to improve model alignment and safety performance.

Why It Might Be a Fit

Must have excellent written English, critical thinking, and analytical reasoning skills, and the ability to consistently evaluate nuanced and policy-sensitive scenarios.

Requirements

  • Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline.
  • 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field.
  • Excellent written English, critical thinking, and analytical reasoning skills.
  • Ability to consistently evaluate nuanced and policy-sensitive scenarios.

Benefits

  • Competitive hourly rate of $60–$70/hour
]]>