Role OverviewThe AI Safety Expert - Red Team will perform jailbreaks, prompt injections, and bias exploitation to conversational AI models and agents, generating high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
What You Will Do
The role involves working independently and asynchronously to meet deadlines while improving AI model performance, documenting reproducibly by producing reports, datasets, and attack cases that customers can act on.
Why It Might Be a Fit
The ideal candidate will have prior experience in red teaming (AI adversarial work, cybersecurity, socio-technical probing), ability to explain risks clearly to technical and non-technical stakeholders, and experience in Adversarial ML, Cybersecurity, and Socio-technical risk.
Requirements
- Fluent in English and Dutch
- Prior experience in red teaming (AI adversarial work, cybersecurity, socio-technical probing)
- Ability to explain risks clearly to technical and non-technical stakeholders
]]>