VideoASMR-Bench: Can AI-Generated ASMR Videos Fool VLMs and Humans?
VideoASMR-Bench: AI生成的ASMR视频能否欺骗视觉语言模型和人类?
Jiaqi Wang, Weijia Wu, Yi Zhan, Rui Zhao, Ming Hu, James Cheng, Wei Liu, Philip Torr, Kevin Qinghong Lin
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
National University of Singapore(新加坡国立大学)
;
Peking University(北京大学)
;
Monash University(墨尔本大学)
;
Video Rebirth
;
University of Oxford(牛津大学)
FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction
FreeSpec: 通过奇异谱重建实现无训练长视频生成
Fangda Chen, Shanshan Zhao, Longrong Yang, Chuanfu Xu, Zhigang Luo, Long Lan
机构
*
College of Computer Science and Technology, National University of Defense Technology(国防科技大学计算机科学与技术学院)
;
Alibaba International Digital Commerce(阿里巴巴国际数字商务)
;
Zhejiang University(浙江大学)
;
Xiangjiang Laboratory(湘江实验室)
Relit-LiVE: Relight Video by Jointly Learning Environment Video
Relit-LiVE: 通过联合学习环境视频实现视频照明
Weiqing Xiao, Hong Li, Xiuyu Yang, Houyuan Chen, Wenyi Li, Tianqi Liu, Shaocong Xu, Chongjie Ye, Hao Zhao, Beibei Wang
机构
*
Nanjing University(南京大学)
;
Tsinghua University(清华大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
Reconstruction or Semantics? What Makes a Latent Space Useful for Robotic World Models
重建还是语义?什么使潜在空间对机器人世界模型有用
Nilaksh, Saurav Jha, Artem Zholus, Sarath Chandar
机构
*
Chandar Research Lab(昌达尔研究实验室)
;
Mila – Quebec AI Institute(魁北克人工智能研究院)
;
Polytechnique Montréal(蒙特利尔理工学院)
;
Canada CIFAR AI Chair(加拿大CIFAR人工智能主席)
UniE2F: A Unified Diffusion Framework for Event-to-Frame Reconstruction with Video Foundation Models
UniE2F: 一种基于视频基础模型的统一扩散框架用于事件到帧重建
Gang Xu, Zhiyu Zhu, Junhui Hou
机构
*
Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系)
;
Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东人工智能与数字经济实验室(深圳))
;
Department of Computer Science, City University of Hong Kong (Dongguan)(香港城市大学(东莞)计算机科学系)
EA-WM: Event-Aware Generative World Model with Structured Kinematic-to-Visual Action Fields
EA-WM:事件感知生成世界模型与结构运动-视觉动作场
Zhaoyang Yang, Yurun Jin, Lizhe Qi, Cong Huang, Kai Chen
机构
*
Fudan University(复旦大学)
;
Zhongguancun Academy(中关村学院)
;
Zhongguancun Institute of Artificial Intelligence(中关村人工智能研究院)
;
University of Science and Technology of China(中国科学技术大学)
;
DeepCybo(深瞳)
机构
*
Gansu Provincial Key Laboratory of Wearable Computing, School of Information Science and Engineering, Lanzhou University(甘肃省可穿戴计算重点实验室,兰州大学信息科学与工程学院)
;
Guangdong-Hong Kong-Macao Joint Laboratory for Emotional Intelligence and Pervasive Computing, Shenzhen MSU-BIT University(粤港澳大湾区情感智能与泛在计算联合实验室,深圳MSU-BIT大学)
;
Department of Computer and Information Engineering, Khalifa University(计算机与信息工程系,哈利法大学)
;
Artificial Intelligence Research Institute, Shenzhen MSU-BIT University(人工智能研究院,深圳MSU-BIT大学)
;
Department of Electrical and Computer Engineering, The University of Hong Kong(电子与计算机工程系,香港大学)