CAP: Controllable Alignment Prompting for Unlearning in LLMs
CAP:用于大语言模型中去学习的可控对齐提示
机构 * School of Information and Software Engineering, University of Electronic Science and Technology of China(电子科技大学信息与软件学院)
专题命中 其他LLM :prompting(title,abstract);LLM(abstract,abstract_cn);large language model(abstract);language model(abstract)
AI总结 本文提出CAP框架,通过强化学习将去学习过程转化为可学习的提示优化,实现可控的去学习,无需更新模型参数,解决了现有方法的计算成本高、遗忘边界不可控等问题。
Comments Accpeted to ACL 2026 Main Conference