BadSKP: Backdoor Attacks on Knowledge Graph-Enhanced LLMs with Soft Prompts
BadSKP: 针对增强知识图谱的大型语言模型的后门攻击
机构 * Ministry of Education Key Lab for Intelligent Networks and Network Security(教育部长智能网络与网络安全重点实验室) ; Xi’an Jiaotong University(西安交通大学) ; INRIA(法国国家信息与自动化技术研究院) ; CFAR, A*STAR(新加坡A*STAR机构) ; Beijing Key Laboratory of Security and Privacy in Intelligent Transportation(北京智能交通安全与隐私重点实验室) ; Beijing Jiaotong University(北京交通大学) ; Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) ; School of Cyber Engineering, Xi’an University of Electronic Science and Technology(西安电子科技大学网络安全工程学院) ; Ministry of Education Key Lab for Intelligent Networks and Network Security at Xi’an Jiaotong University(西安交通大学教育部长智能网络与网络安全重点实验室)
AI总结 本文研究了增强知识图谱的大型语言模型中软提示通道的后门攻击问题,提出BadSKP攻击方法,通过多阶段优化策略有效攻击图到提示接口,实验表明其在冻结和毒化设置下具有高成功率。