CTIGuardian: A Few-Shot Framework for Mitigating Privacy Leakage in Fine-Tuned LLMs
CTIGuardian:一种缓解微调大语言模型隐私泄露的少样本框架
专题命中 隐私与版权 :alignment(abstract);safety(abstract);分类 cs.AI、cs.LG
AI总结 CTIGuardian通过少样本监督整合隐私分类器和红员,提升微调大语言模型的隐私保护与效用平衡。
Comments Accepted at the 18th Cybersecurity Experimentation and Test Workshop (CSET), in conjunction with ACSAC 2025
Journal ref 2025 Annual Computer Security Applications Conference Workshops (ACSAC Workshops), Honolulu, HI, USA, 2025, pp. 510-522