AI Safety Specialist - Fully Remote

Mercor
San Francisco, CA
Job Description
Role Overview

Mercor connects elite creative and technical talent with leading AI research labs. As an AI Safety Specialist, you will be responsible for red teaming conversational AI models and agents, generating high-quality human data, and documenting reproducibly to produce reports, datasets, and attack cases.

What You Will Do

Your day-to-day responsibilities will include identifying jailbreaks, prompt injections, and misuse cases, annotating failures, classifying vulnerabilities, and flagging systemic risks.

Why It Might Be a Fit

You will be a fit for this role if you have prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing, and can explain risks clearly to technical and non-technical stakeholders.

Requirements

  • Fluent Language Skills Required: English & Malay
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing
  • Ability to explain risks clearly to technical and non-technical stakeholders
]]>