Role OverviewAs an AI Safety Red Teamer Expert, you will design adversarial prompts to stress-test frontier AI models, identify vulnerabilities, and contribute to safety benchmarking and red-teaming reports. You will collaborate with AI researchers to improve model alignment, robustness, and safety.
What You Will Do
Your main day-to-day responsibilities will include designing adversarial prompts, evaluating model robustness, documenting vulnerabilities, and contributing to safety reports. You will also collaborate with AI researchers to improve model safety.
Why It Might Be a Fit
To be a fit for this role, you will need strong analytical reasoning, prompt design, and written communication skills. Experience designing adversarial prompts or evaluating frontier AI systems is also required. Preferred qualifications include experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
Requirements
- Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline
- 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field
- Strong analytical reasoning, prompt design, and written communication skills
- Experience designing adversarial prompts or evaluating frontier AI systems
Benefits
]]>