跳到正文
arXiv cs.CV· David Dobre, Leo Schwinn, Gauthier Gidel, Spandana Gella, Perouz Taslakian, Pierre-Andr\'e No\"el·· 4 小时前AI 评分37

视觉记忆注入攻击可经 KV Cache 持续影响模型

Visual Memory Attacks Can Persist Through The KV Cache

AI 导读

研究发现对抗性图像植入的隐藏后门可经 KV cache 持续影响模型,即使图像被移出上下文仍有效。作者提出 Persistent Visual Memory Injection(P-VMI),在 Qwen3-VL-8B-Instruct 上最强配置目标成功率约 90%,且仅在首轮暴露图像时仍有效。

来源:arXiv cs.CV · arxiv.org