ASIDE: Architectural Separation of Instructions and Data in Language Models
ASIDE: 语言模型中指令与数据的架构分离
机构 * Institute of Science and Technology Austria (ISTA)(奥地利科学与技术研究所) ; Fraunhofer Heinrich Hertz Institute(弗劳恩霍夫海因里希·赫兹研究所) ; ELLIS Institute Tübingen(图宾根ELLIS研究所) ; Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) ; Tübingen AI Center(图宾根人工智能中心) ; Centre of eXplainable Artificial Intelligence(可解释人工智能中心) ; Technische Universität Berlin(柏林技术大学) ; Berlin Institute for the Foundations of Learning and Data (BIFOLD)(柏林学习与数据基础研究所)
专题命中 安全训练 :safety(abstract);prompt injection(abstract);分类 cs.LG
AI总结 ASIDE通过在令牌嵌入层面实现指令与数据的分离,提升了语言模型的安全性和性能
Comments ICLR 2026 paper