AI Safety Red Teamer Expert

Mercor
Any Location, IA
Remote
Job Description
Role Overview

Design adversarial prompts to stress-test frontier AI models, identify vulnerabilities, and contribute to safety benchmarking and red-teaming reports. Collaborate with AI researchers to improve model alignment, robustness, and safety.

What You Will Do

Design adversarial prompts, identify jailbreaks, evaluate model robustness, document vulnerabilities, and contribute to safety benchmarking and red-teaming reports.

Why It Might Be a Fit

Must have strong analytical reasoning, prompt design, and written communication skills. Experience designing adversarial prompts or evaluating frontier AI systems is preferred.

Requirements

  • Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field
  • Strong analytical reasoning, prompt design, and written communication skills
  • Experience designing adversarial prompts or evaluating frontier AI systems

Benefits

  • $70–$84/hour
]]>