AI Agents May Always Fall for Prompt Injections
AI Agents May Always Fall for Prompt Injections
机构 * ELLIS Institute Tübingen & MPI-IS & Tübingen AI Center(图宾根ELLIS研究所及MPI-IS与图宾根人工智能中心) ; University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)
专题命中 提示注入 :prompt injection(title,title_cn);alignment(abstract);分类 cs.CL、cs.CY
AI总结 本文基于上下文完整性理论重新审视提示注入问题,揭示了现有防御机制的不足,并提出了一种新的评估框架来设计更安全的自主代理。