SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment
SafeHarness:面向LLM代理部署的生命周期集成安全架构
机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) ; Academy of Mathematics and Systems Science, Chinese Academy of Sciences(中国科学院数学与系统科学研究院) ; Institute of Applied Physics and Computational Mathematics(应用物理与计算数学研究所) ; Peking University(北京大学)
专题命中 记忆与上下文管理 :agent(title,abstract);tool use(abstract);分类 cs.AI
AI总结 本文提出SafeHarness架构,通过四个防御层集成到代理生命周期中,解决现有安全方法的结构性不匹配问题,降低不安全行为和攻击成功率。
Comments 26 pages, 6 figures