DisarmRAG: Stealthy Retriever-Centric Poisoning to Disable Self-Correction in Retrieval-Augmented Generation (Extended Version)
DisarmRAG:以检索器为中心的隐秘中毒攻击,以禁用检索增强生成中的自我纠正(扩展版)
专题命中 检索器与排序 :RAG(summary_cn,abstract);retrieval-augmented generation(title,abstract);retriever(title,abstract);分类 cs.CL
AI总结 研究针对实际部署中LLMs自我纠正能力削弱RAG攻击效果的问题,提出DisarmRAG这一专注检索器的新型中毒范式,构建含迭代协同优化与对比学习技术的攻击框架,经多模型和基准测试验证其有效性及隐秘性。
Comments This paper is an extended version of our original paper accepted by ACM CCS 2026