Role OverviewAI Safety Expert - Red Team will be responsible for conversational AI models and agents, focusing on jailbreaks, prompt injections, misuse cases, and bias exploitation. The role involves generating high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
What You Will Do
The main day-to-day responsibilities of this role include red team conversational AI models and agents, generating human data, applying structure, documenting reproducibly, uncovering vulnerabilities, and delivering reproducible artifacts.
Why It Might Be a Fit
To be a fit for this role, you should have prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing, and the ability to explain risks clearly to both technical and non-technical stakeholders.
Requirements
- Fluent in English and Assamese.
- Prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing.
- Ability to explain risks clearly to both technical and non-technical stakeholders.
- Adaptability to move across projects and customers.
]]>