Malicious and Unintentional Disclosure Risks in Large Language Models for Code Generation
大型语言模型在代码生成中的恶意与非故意披露风险
机构 * Digital Safety Research Institute(数字安全研究研究院) ; UL Research Institutes(UL研究机构)
专题命中 预训练与数据 :language model(title,summary_cn);large language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.LG
AI总结 本文研究了大型语言模型在代码生成训练中可能泄露训练数据中的敏感信息的风险,分析了非故意和恶意披露两种风险,并通过Open Language Model模型评估了不同数据集和模型版本的风险变化。
Comments The 3rd International Workshop on Mining Software Repositories Applications for Privacy and Security (MSR4P&S), co-located with SANER 2025