AI Safety Red Teamer Expert

Mercor
San Francisco, CA
Job Description
Role Overview

As an AI Safety Red Teamer Expert, you will design adversarial prompts to stress-test frontier AI models, identify unsafe behaviors, and evaluate model robustness across various sensitive domains. You will collaborate with AI researchers to improve model alignment, robustness, and safety. This is a contract position.

What You Will Do

Design adversarial prompts, identify unsafe behaviors, evaluate model robustness, document vulnerabilities, and contribute to safety benchmarking and red-teaming reports.

Why It Might Be a Fit

You will be a fit for this role if you have strong analytical reasoning, prompt design, and written communication skills, and experience designing adversarial prompts or evaluating frontier AI systems.

Requirements

  • Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Benefits

  • $70–$84/hour
]]>