Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations
LLMs是否了解工具无关性?解密工具调用中的结构对齐偏差
机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) ; School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院) ; Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; School of Computer Science and Technology, Donghua University(东华大学计算机科学与技术学院)
专题命中 AI治理与伦理 :alignment(title,abstract);分类 cs.CL、cs.AI
AI总结 本文研究LLMs在面对无关工具时的调用偏差,提出SABEval数据集和Contrastive Attention Attribution方法,揭示结构对齐偏差的成因并提出缓解策略。
Comments Accepted to ACL 2026 (Main Conference)