AI Safety Expert - Red Team

Mercor
San Francisco, CA
Job Description
Role Overview

As an AI Safety Expert, you will be part of Mercor's red team, responsible for conversational AI models and agents to identify vulnerabilities. You will work independently and asynchronously to meet deadlines while improving AI model performance.

What You Will Do

Your main responsibilities will include red teaming, annotating failures, classifying vulnerabilities, flagging systemic risks, and documenting findings reproducibly to create reports, datasets, and attack cases for customer action.

Why It Might Be a Fit

To be a fit for this role, you should have strong communication skills to explain risks to both technical and non-technical stakeholders, and experience in red teaming, AI adversarial work, or cybersecurity.

Requirements

  • Fluent in English and Assamese
  • Prior experience in red teaming, AI adversarial work, or cybersecurity
  • Strong communication skills to explain risks to both technical and non-technical stakeholders
]]>