ChainCaps: Composition-Safe Tool-Using Agents via Monotonic Capability Attenuation
ChainCaps: 通过单调能力衰减实现组合安全的工具使用智能体
机构 * Independent Researcher, Seattle, WA, USA(华盛顿州塞勒姆独立研究员) ; Independent Researcher, New York City, NY, USA(纽约市纽约独立研究员) ; King Abdullah University of Science and Technology(国王阿卜杜勒阿齐兹科学技术大学)
专题命中 安全训练 :safety(abstract,comments);分类 cs.AI
AI总结 针对工具组合中的权限洗钱漏洞,提出ChainCaps机制,通过运行时能力预算交集传播规则,在不修改智能体或工具服务器的情况下,将攻击成功率从25-68%降至0-4.8%,同时保持96-100%的良性任务完成率。
Comments Published at the Second Workshop on Agents in the Wild: Safety, Security, and Beyond (AIWILD) at ICML 2026