Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents
智能体技能可能有害:LLM智能体中技能诱导故障的实证研究
机构 * Huazhong University of Science and Technology(华中科技大学) ; Microsoft Research(微软研究院) ; Microsoft(微软) ; University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
AI总结 本文通过提出差异分析框架,在SkillsBench等数据集上发现LLM智能体的技能会引发功能故障与效率回归,并构建SkillTriage工具,为安全复用技能提供研究方向。