Key Vulnerabilities Uncovered by AI Red Teaming

AI red teaming aims to expose prompt injection, jailbreaks, data exfiltration, adversarial attacks, model bias, and excessive agency, which can lead to harmful actions or data exposure.

Sources

Open the full topic