VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders
VideoRAE:通过表示自动编码器驯服用于生成建模的视频基础模型
Zhihao Xie, Junfeng Wu, Xinting Hu, Junchao Huang, Li Jiang
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Huazhong University of Science and Technology(华中科技大学)
;
Shenzhen Loop Area Institute(深圳河套学院)
;
University of Science and Technology of China(中国科学技术大学)
FlowWAM: Optical Flow as a Unified Action Representation for World Action Models
FlowWAM:光流作为世界动作模型的统一动作表示
Yixiang Chen, Peiyan Li, Yuan Xu, Qisen Ma, Jiabing Yang, Kai Wang, Jianhua Yang, Dong An, He Guan, Gaoteng Liu, Jianlou Si, Jun Huang, Jing Liu, Nianfeng Liu, Yan Huang, Liang Wang
机构
*
New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
FiveAges(无)
;
MBZUAI(无)
;
Alibaba Group(阿里巴巴集团)
机构
*
School of Cyber Science and Engineering, Huazhong University of Science and Technology(华中科技大学网络空间安全学院)
;
College of Computer Science, Chongqing University(重庆大学计算机科学学院)
;
School of Software and engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院)
;
School of Information and Communication Technology, Griffith University(格里菲斯大学信息与通信技术学院)
Progression as Latent Drift: Generative Forecasting of Slow-Evolving Pathologies
作为潜在漂移的进展:缓慢演变病理的生成预测
Yuxiang Feng, Juncheng Wang, Chao Xu, Wenlong Hou, Huihan Wang, Yijie Qian, Yang Liu, Baigui Sun, Yong Liu, Shujun Wan
机构
*
Zhejiang University(浙江大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
IROOTECH TECHNOLOGY(IROOTECH技术)
;
Wolf 1069 b Lab, Sany Group(Wolf 1069 b实验室,三一集团)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Fudan University(复旦大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Shanghai Jiao Tong University(上海交通大学)
Straight-Path Flow Matching for Incomplete Multi-View Clustering
用于不完整多视图聚类的直线路径流匹配
Yiteng Yuan, Junyan Wang, Zheyuan Liu, Hong Jia, Lei Fan, Zhulin Tao, Lianbo Guo
机构
*
School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院)
;
Australian Institute for Machine Learning, Adelaide University(澳大利亚机器学习研究所,阿德莱德大学)
;
University of Auckland(奥克兰大学)
;
University of New South Wales(新南威尔士大学)
;
Communication University of China(中国通信大学)
Multiplayer Interactive World Models with Representation Autoencoders
基于表示自编码器的多智能体交互世界模型
Anthony Hu, Václav Volhejn, Adrien Ramanana Rahary, Chris Mulder, Aditya Makkar, Alyx Liao, Amélie Royer, Manu Orsini, Adam Jelley, Eloi Alonso, Florian Laurent, Fredrik Norén, James Swingos, Jan Hünermann, Kent Rollins, Lucas Hosseini, Matthieu Le Cauchois, Maxim Peter, Pim de Witte, Tim Brown, Vincent Micheli, Moritz Böhle, Gabriel de Marmiesse, Viktoriia Sharmanska, Lucia Specia, Michael Black, Patrick Pérez
机构
*
Epic Games
;
École nationale des ponts et chaussées(法国国家桥梁与道路学院)