Making MLLMs Blind: Adversarial Smuggling Attacks in MLLM Content Moderation
让 MLLMs 失明:MLLM 内容审核中的对抗走私攻击
机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, CASIA(中国科学院自动化研究所多模态人工智能系统国家重点实验室) ; School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) ; Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京市多模态信息超智能安全重点实验室) ; Hellogroup ; University of Washington(华盛顿大学) ; Jilin University(吉林大学) ; ShanghaiTech University(上海科技大学)
专题命中 幻觉与鲁棒性 :MLLM(title);multimodal large language model(abstract);分类 cs.CV
AI总结 本文揭示了MLLM在内容审核中面临的新威胁——对抗走私攻击,通过SmuggleBench基准测试发现多种模型易受攻击,提出通过感知和推理角度分析其根本原因并探索缓解策略。
Comments Accepted to ACL 2026. 19 pages, 6 figures