MosaicMem: Hybrid Spatial Memory for Controllable Video World Models
MosaicMem: 用于可控视频世界模型的混合空间记忆
Wei Yu, Runjia Qian, Yumeng Li, Liquan Wang, Songheng Yin, Sri Siddarth Chakaravarthy P, Dennis Anthony, Yang Ye, Yidi Li, Weiwei Wan, Animesh Garg
机构
*
University of Toronto(多伦多大学)
;
Vector Institute(向量研究所)
;
The University of Osaka(大阪大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Mujin Inc.(Mujin公司)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Taiyuan University of Technology(太原本科技大学)
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
LimX Dynamics Technology Co., Ltd.(LimX动力技术有限公司)
;
Shandong University(山东大学)
;
Institute of Deep Perception Technology, Jiangsu Industrial Technology Research Institute (JITRI)(江苏工业技术研究院深度感知技术研究所)
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
S-Lab, Nanyang Technological University(南洋理工大学S实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
Fudan University(复旦大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Adobe Research(Adobe研究实验室)
;
Stanford University(斯坦福大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
CPII under InnoHK Project(InnoHK项目下的CPII)
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Faculty of Engineering and IT, University of Technology Sydney(悉尼大学工程与信息学院)
;
Xiamen University(厦门大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
CommentsThe authors have decided to withdraw this article due to the following reasons identified after publication: Experimental Errors: Significant inaccuracies were discovered in the experimental results concerning segmentation and depth estimation. Authorship Disputes: In addition to the technical issues, there are unresolved disagreements regarding the author sequence and contributions