CommentsThe paper is withdrawn due to the need for further revision and verification of experimental results. A revised version will be resubmitted once the updates are completed
机构
*
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机学院,北京大学)
;
Huawei Inc.(华为公司)
Uniform Discrete Diffusion with Metric Path for Video Generation
Haoge Deng, Ting Pan, Fan Zhang, Yang Liu, Zhuoyan Luo, Yufeng Cui, Wenxuan Wang, Chunhua Shen, Shiguang Shan, Zhaoxiang Zhang, Xinlong Wang
机构
*
National Laboratory of Pattern Recognition, CASIA(中国科学院自动化所模式识别国家实验室)
;
Key Laboratory of Intelligent Information Processing, ICT, CAS(中国科学院信息科技重点实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Zhejiang University(浙江大学)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
From Seeing to Predicting: A Vision-Language Framework for Trajectory Forecasting and Controlled Video Generation
Fan Yang, Zhiyang Chen, Yousong Zhu, Xin Li, Jinqiao Wang
机构
*
Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(基础模型研究中心、自动化研究所、中国科学院)
;
Peng Cheng Laboratory, Shenzhen, China(鹏城实验室、深圳中国)
;
School of Artificial Intelligence, University of Chinese Academy of Science, Beijing, China(人工智能学院、中国科学院大学、北京中国)
;
Wuhan AI Research, Wuhan, China(武汉人工智能研究、武汉中国)
;
MAPLE Lab, Westlake University(MAPLE实验室、西湖大学)