FoE: Forest of Errors Makes the First Solution the Best in Large Reasoning Models
FoE:错误森林使大推理模型中的首个解成为最佳解
机构 * School of Software and Microelectronics, Peking University(北京大学软件与微电子学院) ; State Key Laboratory of General Artificial Intelligence, School of Intelligence Science and Technology, Peking University(北京大学智能科学与技术学院通用人工智能国家重点实验室)
专题命中 测试时计算 :reasoning(title,abstract);分类 cs.CL、cs.AI
AI总结 本文发现大推理模型中首个解往往最优,提出RED框架通过抑制错误增长和修剪后续错误提升性能,实验显示在多个基准上性能提升达19.0%。