AI Safety Expert - Red Team

Mercor
San Francisco, CA
Job Description
Role Overview

As an AI Safety Expert, you will work on a contract basis to identify vulnerabilities in conversational AI models and agents. You will be responsible for generating high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks. You will work independently and asynchronously to meet deadlines while improving AI model performance.

What You Will Do

Your main responsibilities will include red teaming conversational AI models and agents, generating human data, applying structure to maintain consistent testing, and documenting reproducibly by producing reports, datasets, and attack cases.

Why It Might Be a Fit

To be a good fit for this role, you should have prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing, as well as the ability to communicate risks clearly to technical and non-technical stakeholders.

Requirements

  • Fluent Language Skills: Native fluency in English & Thai.
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
  • Ability to communicate risks clearly to technical and non-technical stakeholders.
]]>