Role OverviewThe AI Safety Expert - Red Team will be responsible for identifying jailbreaks, prompt injections, and misuse cases in conversational AI models and agents. They will generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks. The expert will apply structure by following taxonomies, benchmarks, and playbooks to ensure consistent testing, and document reproducibly to produce reports, datasets, and attack cases that customers can act on.
What You Will Do
The main day-to-day responsibilities of the AI Safety Expert - Red Team will include conversational AI testing, identifying and mitigating risks, and generating high-quality human data. They will also review AI outputs on sensitive topics such as bias and misinformation.
Why It Might Be a Fit
The ideal candidate will have prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing, and be able to explain risks clearly to technical and non-technical stakeholders. Experience in Adversarial ML, background in Cybersecurity, and expertise in socio-technical risk are preferred qualifications.
Requirements
- Fluent Language Skills Required: English & Malay
- Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing
- Ability to explain risks clearly to technical and non-technical stakeholders
]]>