Pretraining Data Exposure in Large Language Models: A Survey of Membership Inference, Data Contamination, and Security Implications
大型语言模型中的预训练数据暴露:成员推断、数据污染及安全影响综述
机构 * Japan Advanced Institute of Science and Technology(日本先进科学研究院)
专题命中 预训练与数据 :large language model(title,abstract);language model(title,abstract);pretraining(title,abstract);LLM(abstract,abstract_cn)
AI总结 本文首次统一综述了大型语言模型中的预训练数据暴露问题,涵盖成员推断和数据污染,形式化定义了暴露级别,回顾了攻击与防御方法,并总结了实证发现及未来研究方向。
Comments accepted by NLDB 2025