Role OverviewEvaluate AI-generated responses for safety, factual accuracy, policy compliance, and quality. Review content involving sensitive domains and apply evaluation rubrics for AI safety benchmarking.
What You Will Do
Evaluate AI-generated responses, identify unsafe outputs, and provide feedback to improve model alignment and safety performance.
Why It Might Be a Fit
Must have excellent written English, critical thinking, and analytical reasoning skills, with ability to evaluate nuanced and policy-sensitive scenarios.
Requirements
- Bachelor's degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related discipline.
- 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field.
- Excellent written English, critical thinking, and analytical reasoning skills.
- Ability to consistently evaluate nuanced and policy-sensitive scenarios.
]]>