AI Safety Expert - Red Team

Mercor
New York, NY
Job Description
Role Overview

The AI Safety Expert - Red Team will be responsible for conversational AI models and agents, focusing on jailbreaks, prompt injections, misuse cases, and bias exploitation. The role requires generating high-quality human data, annotating failures, classifying vulnerabilities, and flagging systemic risks. The expert will apply structure using taxonomies, benchmarks, and playbooks to maintain consistent testing and document reproducibly.

What You Will Do

The main day-to-day responsibilities of the AI Safety Expert - Red Team include conversational AI model testing, generating human data, annotating failures, and classifying vulnerabilities. The expert will also produce reports, datasets, and attack cases that customers can act on.

Why It Might Be a Fit

To be a fit for this role, you should have prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing. You should also be able to communicate risks clearly to both technical and non-technical stakeholders and be adaptable to move across projects and customers.

Requirements

  • Fluent Language Skills Required: Native fluency in English & Punjabi.
  • Prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing.
  • Ability to communicate risks clearly to both technical and non-technical stakeholders.
  • Adaptability to move across projects and customers.
]]>