Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning
通过强化学习提升多图像接地在MLLMs中的推理能力
Bob Zhang, Haoran Li, Tao Zhang, Jianan Li, Cilin Yan, Xikai Liu, Jiayin Cai, Yanbin Hao
机构
*
Xiaohongshu Inc.(小红书公司)
;
University of Science and Technology of China(中国科学技术大学)
;
Wuhan University(武汉大学)
;
Technical University of Munich(慕尼黑工业大学)
;
Hefei University of Technology(合肥工业大学)
机构
*
Xiaohongshu Inc.(小红书公司)
;
College of Control Science and Technology, Zhejiang University(浙江大学控制科学与工程学院)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
School of Computing and Artificial Intelligence, Shanghai University of Finance and Economics(上海财经大学计算与人工智能学院)
;
State Key Lab of General AI, School of Intelligence Science and Technology, Peking University(北京大学智能科学与技术学院通用人工智能国家重点实验室)
;
Squirrel AI(松鼠AI)
;
Center for Data Science, Peking University(北京大学数据科学中心)
;
Institute for Artificial Intelligence, Peking University(北京大学人工智能研究院)