The Illusion of Diminishing Returns: Measuring Long Horizon Execution in LLMs
持续扩展的幻觉回报:衡量LLM的长周期执行
机构 * University of Cambridge(剑桥大学) ; Institute for AI, University of Stuttgart(斯图加特大学人工智能研究所) ; Max Planck Institute for Intelligent Systems(智能系统马克斯·普朗克研究所) ; ELLIS Institute Tübingen(图宾根ELLIS研究所) ; University of Southampton(南安普顿大学) ; Tübingen AI Center(图宾根人工智能中心)
专题命中 测试时计算 :reasoning(abstract);test-time compute(abstract);分类 cs.AI
AI总结 本文通过分析LLM在长周期任务中的执行能力,揭示了持续扩展LLM并非必然导致回报递减,而是执行能力的提升能显著改善长周期任务表现。
Comments Published at ICLR 2026