Caught in the Act(ivation): Toward Pre-Output and Multi-Turn Detection of Credential Exfiltration by LLM Agents
当场抓获(激活):面向LLM智能体的凭证泄露预输出和多轮检测
机构 * University of California, Berkeley(加州大学伯克利分校)
专题命中 提示注入 :prompt injection(abstract);分类 cs.AI
AI总结 研究通过激活探针、蜜令令牌和累积信息流追踪三种互补防御方法,在预输出和多轮对话中检测LLM智能体的凭证泄露。