Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models
理解强化学习在多模态推理模型后训练中幻觉的作用
机构 * University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) ; University of Science and Technology of China(中国科学技术大学) ; Arizona State University(亚利桑那州立大学) ; Honda Research Institute, USA(本田美国研究所)
专题命中 视觉推理 :visual reasoning(abstract);multimodal large language model(abstract);MLLM(abstract);分类 cs.CV、cs.AI、cs.LG
AI总结 本文提出Hallucination-as-Cue框架,通过引入模态特异性扰动分析强化学习对多模态推理模型的影响,揭示幻觉在训练中的关键作用,挑战现有假设。
Comments CVPR 2026