SrDetection: A Self-Referential Framework for Data Leakage Detection in Code Large Language Models
SrDetection: 一种用于代码大型语言模型中数据泄露检测的自参考框架
机构 * Shenzhen Key Laboratory for High Performance Data Mining(深圳高性能数据挖掘重点实验室) ; Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究院,中国科学院) ; Shenzhen University(深圳大学) ; University of Science and Technology of China(中国科学技术大学) ; PolyU(香港理工大学) ; East China Normal University(华东师范大学) ; Artificial Intelligence Research Institute, Shenzhen University of Advanced Technology(深圳大学先进技术研究院人工智能研究所) ; University of New South Wales(新南威尔士大学) ; Anhui University(安徽大学)
AI总结 提出自参考框架SrDetection,通过对比模型对原始样本与语义等价变体的行为差异检测数据泄露,在灰盒和黑盒设置下均优于基线方法。