Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models
理解强化学习在多模态推理模型后训练中幻觉的作用
机构 * University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) ; University of Science and Technology of China(中国科学技术大学) ; Arizona State University(亚利桑那州立大学) ; Honda Research Institute, USA(本田美国研究所)
AI总结 本文提出Hallucination-as-Cue框架,通过引入模态特异性扰动分析强化学习对多模态推理模型的影响,揭示幻觉在训练中的关键作用,挑战现有假设。
Comments CVPR 2026