AI Safety Specialist

Mercor
New York, NY
Job Description
Role Overview

As an AI Safety Specialist, you will be responsible for red teaming conversational AI models and agents, generating high-quality human data, and documenting reproducibly. You will work on jailbreaks, prompt injections, misuse cases, and bias exploitation, and communicate risks clearly to both technical and non-technical stakeholders.

What You Will Do

Your day-to-day responsibilities will include red teaming conversational AI models and agents, generating high-quality human data, annotating failures, classifying vulnerabilities, and flagging systemic risks.

Why It Might Be a Fit

This role requires strong communication skills, adaptability, and the ability to move across projects and customers. Experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing is a must-have, and experience with Adversarial ML, Cybersecurity, or Socio-technical risk is preferred.

Requirements

  • Fluent Language Skills Required: Native fluency in English & Punjabi.
  • Prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing.
  • Ability to communicate risks clearly to both technical and non-technical stakeholders.
  • Adaptability to move across projects and customers.
]]>