AI · 126 giorni fa
Un modello che appare resiliente in un test one-shot può comunque essere facile da indirizzare in una conversazione reale. L'errore standard di procurement è trattare la sicurezza a singolo turno come un proxy di come si comporta un chatbot quando un attacker può continuare ad adattarsi tra i turni.
3 fonti che coprono questa storia
Frontier AI models collapse under multi-turn AI attacks, Cisco finds - Help Net Security
Cisco research finds multi-turn AI attacks push success rates as high as 88% across 15 flagship models from OpenAI, Anthropic, Google, xAI.
AI models more vulnerable than claimed when faced with iterative attacks
Cisco researchers show how leading AI models wither under realistic multi-turn attacks, calling into question the value of vendors’ single-prompt safety benchmarks.
Leading AI models are more vulnerable to malicious prompts than vendors claim
Hackers could subvert frontier models with attacks that their developers overlook, Cisco said.
Part of the PlainSec briefing for 2026-05-28