VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders
VideoRAE:通过表示自动编码器驯服用于生成建模的视频基础模型
Zhihao Xie, Junfeng Wu, Xinting Hu, Junchao Huang, Li Jiang
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Huazhong University of Science and Technology(华中科技大学)
;
Shenzhen Loop Area Institute(深圳河套学院)
;
University of Science and Technology of China(中国科学技术大学)
FlowWAM: Optical Flow as a Unified Action Representation for World Action Models
FlowWAM:光流作为世界动作模型的统一动作表示
Yixiang Chen, Peiyan Li, Yuan Xu, Qisen Ma, Jiabing Yang, Kai Wang, Jianhua Yang, Dong An, He Guan, Gaoteng Liu, Jianlou Si, Jun Huang, Jing Liu, Nianfeng Liu, Yan Huang, Liang Wang
机构
*
New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
FiveAges(无)
;
MBZUAI(无)
;
Alibaba Group(阿里巴巴集团)
机构
*
School of Cyber Science and Engineering, Huazhong University of Science and Technology(华中科技大学网络空间安全学院)
;
College of Computer Science, Chongqing University(重庆大学计算机科学学院)
;
School of Software and engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院)
;
School of Information and Communication Technology, Griffith University(格里菲斯大学信息与通信技术学院)
Progression as Latent Drift: Generative Forecasting of Slow-Evolving Pathologies
作为潜在漂移的进展:缓慢演变病理的生成预测
Yuxiang Feng, Juncheng Wang, Chao Xu, Wenlong Hou, Huihan Wang, Yijie Qian, Yang Liu, Baigui Sun, Yong Liu, Shujun Wan
机构
*
Zhejiang University(浙江大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
IROOTECH TECHNOLOGY(IROOTECH技术)
;
Wolf 1069 b Lab, Sany Group(Wolf 1069 b实验室,三一集团)