AI Safety Specialist - Fully Remote | Upto $22/hr

Mercor
San Francisco, CA
Category Healthcare
Job Description
Role Overview

The AI Safety Specialist will work independently and asynchronously to identify jailbreaks, prompt injections, and misuse cases in conversational AI models and agents. They will generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.

What You Will Do

The Specialist will apply structure by following taxonomies, benchmarks, and playbooks to maintain consistent testing, and document reproducibly by producing reports, datasets, and attack cases that customers can act on.

Why It Might Be a Fit

The ideal candidate will have strong judgment about language and content accuracy, rigorous attention to detail and consistency, and the ability to explain reasoning clearly to both technical and non-technical audiences.

Requirements

  • Fluent in English and Telugu
  • Strong judgment about language and content accuracy
  • Rigorous attention to detail and consistency
  • Ability to explain reasoning clearly to both technical and non-technical audiences
  • Adaptability across projects, task types, and customers
]]>