Free-MoRef: Instantly Multiplexing Context Perception Capabilities of Video-MLLMs within Single Inference
机构 * Sun Yat-sen University(中山大学) ; Peng Cheng Laboratory(鹏城实验室) ; OPPO AI Center(OPPO人工智能中心) ; Research Institute(研究 institute) ; The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) ; Harbin Institute of Technology(哈尔滨工业大学) ; Shenzhen Key Laboratory of Digital Living Network and Content Service(深圳数字生活网络与内容服务重点实验室) ; Guangdong Key Laboratory of Big Data Analysis and Processing(广东省大数据分析与处理重点实验室)
专题命中 视频理解 :video understanding(abstract);long video(abstract);分类 cs.CV
Comments published in ICCV 2025