GRACE: Boosting Video MLLMs with Grounded Action-Centric Evidence for Viewer Sentiment Prediction
GRACE: 基于接地动作中心证据增强视频多模态大语言模型用于观众情感预测
Ruoxuan Yang, Tieyuan Chen, Xiaofeng Huang, Haibing Yin, Jun Wang, Xiping Chen, Jun Yin, Xuesong Gao, Weiyao Lin
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Hangzhou Dianzi University(杭州电子科技大学)
;
The 52nd Research Institute of China Electronics Technology Group Corporation(中国电子科技集团公司第五十二研究所)
;
Hangzhou Bywin Technology Co., Ltd.(杭州百威科技有限公司)
;
Zhejiang Dahua Technology Co., Ltd.(浙江大华技术股份有限公司)
;
School of Information Science and Engineering, Shandong University(山东大学信息科学与工程学院)
;
Haihe Laboratory of Information Technology Application Innovation(海河信息技术应用创新实验室)
专题命中
GUI与屏幕智能体
:MLLM(summary_cn,abstract);multimodal large language model(abstract);分类 cs.CV
Comments6 pages, 5 figures. Published in MobiSys Workshop '26
Journal refIn Proceedings of the 24th Annual International Conference on Mobile Systems, Applications and Services Workshops (MobiSys Workshop '26), June 21-25, 2026, Cambridge, United Kingdom. ACM, New York, NY, USA, 6 pages
SAIN: Structure-Aware Interactive Navigation with Active Dialogue Grounding for Mobile Robot
SAIN:面向移动机器人的、结合主动对话 grounding 的结构感知交互式导航
Yuhao Cao, Xiao Liu, Yang Xie, Lu Liu, Haoyao Chen
机构
*
School of Mechanical Engineering and Automation, Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)机械工程与自动化学院)
;
Department of Mechanical Engineering, City University of Hong Kong(香港城市大学机械工程系)
Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation
用于动态视觉语言导航的从慢速推理器到快速规划器的逐令牌潜在流
Tianshuai Hu, Yangyi Zhong, Zeying Gong, Lingdong Kong, Xiaodong Mei, Guoyang Zhao, Xiaolu Liu, Song Wang, Rong Li, Junwei Liang
机构
*
The Hong Kong University of Science and Technology(香港科技大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
National University of Singapore(新加坡国立大学)
;
Zhejiang University(浙江大学)