AI Safety Expert - Red Team

Mercor
New York, NY
Job Description
Role Overview

AI Safety Expert - Red Team will be responsible for conversational AI models and agents, focusing on jailbreaks, prompt injections, misuse cases, and bias exploitation. The role involves generating high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.

What You Will Do

The main day-to-day responsibilities of this role include red team conversational AI models and agents, generating human data, applying structure, documenting reproducibly, uncovering vulnerabilities, and delivering reproducible artifacts.

Why It Might Be a Fit

To be a fit for this role, you should have prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing, and the ability to explain risks clearly to both technical and non-technical stakeholders.

Requirements

  • Fluent in English and Assamese.
  • Prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing.
  • Ability to explain risks clearly to both technical and non-technical stakeholders.
  • Adaptability to move across projects and customers.
]]>