AI Safety Specialist

Mercor
New York, NY
Job Description
Role Overview

The AI Safety Specialist will work independently and asynchronously to identify jailbreaks, prompt injections, and misuse cases in conversational AI models. They will generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks. The Specialist will apply structure by following taxonomies, benchmarks, and playbooks to maintain consistent testing and document reproducibly by producing reports, datasets, and attack cases.

What You Will Do

The main day-to-day responsibilities of the AI Safety Specialist include red teaming conversational AI models, generating human data, and applying structure to maintain consistent testing.

Why It Might Be a Fit

The ideal candidate will have native fluency in English and Odia, prior red teaming experience in AI adversarial work, and strong communication skills to explain risks to technical and non-technical stakeholders.

Requirements

  • Native fluency in English and Odia
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing
  • Strong communication skills

Benefits

  • $20-$22/hour
]]>