AI Safety Specialist - Fully Remote

Mercor
New York, NY
Job Description
Role Overview

Mercor connects elite creative and technical talent with leading AI research labs. As an AI Safety Specialist, you will be responsible for red teaming conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases, and bias exploitation. You will also generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.

What You Will Do

Your main day-to-day responsibilities will include red teaming conversational AI models and agents, generating human data, applying structure to maintain consistent testing, documenting reproducibly, and reviewing AI outputs on sensitive topics.

Why It Might Be a Fit

We are looking for a candidate with a curious and adversarial mindset, strong communication skills, and adaptability to move across projects and customers. Experience with Adversarial ML, Cybersecurity, and socio-technical risk is a plus.

Requirements

  • Fluent Language Skills Required: English & Assamese
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing
  • Ability to push systems to breaking points with a curious and adversarial mindset
  • Structured approach using frameworks or benchmarks
  • Strong communication skills to explain risks to technical and non-technical stakeholders
  • Adaptability to move across projects and customers
]]>