ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning
ThoughtFold: 通过内省偏好学习折叠推理链
机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; University of Science and Technology of China(中国科学技术大学) ; MMLab, The Chinese University of Hong Kong(香港中文大学MMLab)
专题命中 其他推理 :reasoning(title,abstract);CoT(abstract,abstract_cn);分类 cs.AI
AI总结 提出ThoughtFold框架,通过细粒度偏好学习惩罚冗余探索并鼓励直接连接关键推理段,将推理链折叠为更简洁路径,在保持精度的同时大幅降低token使用量。