When Agents Disagree With Themselves: Behavioral Consistency as an Uncertainty Signal for LLM Agents
当智能体与自身意见相左:测量基于LLM的智能体的行为一致性
机构 * Aman Mehta
专题命中 知识编辑与模型理解 :LLM(title,title_cn);分类 cs.AI
AI总结 研究发现基于LLM的智能体在相同任务上运行结果不一致,且这种不一致与任务成功率密切相关,通过监控行为一致性可提升智能体可靠性。
Comments Accepted at the ICML 2026 Workshop on Statistical Frameworks for Uncertainty in Agentic Systems. 12 pages, 9 figures