No AI model is fully resistant to bioweapon queries, Cisco found. Attack success rates hit 88%.
Cisco researchers bypassed safety guardrails on ChatGPT, Claude, and Gemini within five conversational turns, eliciting information about biological weapons by gradually steering conversations around the