机构
*
School of Computer Science and Technology, Tianjin University(天津大学计算机科学与技术学院)
;
School of Computer Science and Technology, Tiangong University(天津工大学计算机科学与技术学院)
;
Key Research Center for Surface Monitoring and Analysis of Relics, State Administration of Cultural Heritage(文物表面监测与分析关键研究中心,国家文物局)
机构
*
Department of Computer Science and Software Engineering, The University of Western Australia(计算机科学与软件工程系,西澳大学)
;
Munich Center for Machine Learning (MCML) and Technical University of Munich (TUM)(慕尼黑机器学习中心(MCML)和技术大学慕尼黑(TUM))
;
School of Information Technology, Murdoch University(信息科技学院,墨尔本大学)
;
Department of Electrical, Electronics and Computer Engineering, The University of Western Australia(电子、电子与计算机工程系,西澳大学)
机构
*
College of Science, Mathematics and Technology, Wenzhou-Kean University(科学、数学与技术学院,温州市凯恩大学)
;
State Key Laboratory of Complex & Critical Software Environment, Beihang University(复杂与关键软件环境国家重点实验室,北航)
;
AI Security Lab(360人工智能安全实验室)
S$^2$Q-VDiT: Accurate Quantized Video Diffusion Transformer with Salient Data and Sparse Token Distillation
S$^2$Q-VDiT: 精确量化视频扩散变换器与显著数据和稀疏令牌蒸馏
Weilun Feng, Haotong Qin, Chuanguang Yang, Xiangqi Li, Han Yang, Yuqi Li, Zhulin An, Libo Huang, Michele Magno, Yongjun Xu
机构
*
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(人工智能安全国家重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
ETH Zürich(苏黎世联邦理工学院)
机构
*
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(人工智能安全国家重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
ETH Zürich(苏黎世联邦理工学院)
;
City College of New York, City University of New York, USA(纽约城市学院,纽约城市大学,美国)
;
Shanghai Jiao Tong University(上海交通大学)
TDM-R1: Reinforcing Few-Step Diffusion Models with Non-Differentiable Reward
TDM-R1: 通过非可微奖励强化少步扩散模型
Yihong Luo, Tianyang Hu, Weijian Luo, Jing Tang
机构
*
Hong Kong University of Science and Technology(香港科技大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
hi-Lab, Xiaohongshu Inc(小红书实验室)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
MTVCraft: Tokenizing 4D Motion for Arbitrary Character Animation
MTVCraft: 4D运动分词用于任意角色动画
Yanbo Ding, Xirui Hu, Zhizhi Guo, Yan Zhang, Xinrui Wang, Zhixiang He, Chi Zhang, Yali Wang, Xuelong Li
机构
*
Shenzhen Key Laboratory of Computer Vision and Pattern Recognition, Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences, Shenzhen, China(深圳计算机视觉与模式识别重点实验室,深圳先进技术研究院,中国科学院,深圳,中国)
;
Institute of Artificial Intelligence (TeleAI), China Telecom, Beijing, China(人工智能研究所(TeleAI),中国电信,北京,中国)
;
School of Computer Science and Technology, Xi’an Jiaotong University, Xi’an, China(计算机科学与技术学院,西安交通大学,西安,中国)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences, Beijing, China(人工智能学院,中国科学院大学,北京,中国)
;
Shanghai Artificial Intelligence Laboratory, Shanghai, China(上海人工智能实验室,上海,中国)
Video-EM: Event-Centric Episodic Memory for Long-Form Video Understanding
视频-EM:面向长视频理解的事件中心型片段记忆
Yun Wang, Long Zhang, Jingren Liu, Jiaqi Yan, Zhanjie Zhang, Jiahao Zheng, Ao Ma, Run Ling, Xun Yang, Dapeng Wu, Xiangyu Chen, Xuelong Li
机构
*
City University of Hong Kong(香港城市大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Tianjin University(天津大学)
;
Nanjing University(南京大学)
;
Zhejiang University(浙江大学)
;
The Institute of Artificial Intelligence (TeleAI), China Telecom(中国电信人工智能研究所(TeleAI))
SurgCUT3R: Surgical Scene-Aware Continuous Understanding of Temporal 3D Representation
SurgCUT3R: 术场景感知的连续时间3D表示理解
Kaiyuan Xu, Fangzhou Hong, Daniel Elson, Baoru Huang
机构
*
The Hamlyn Centre for Robotic Surgery, Imperial College London(机器人手术中心,帝国理工学院伦敦分校)
;
S-Lab, College of Computing and Data Science, Nanyang Technological University(S实验室, computing and Data Science学院,南洋理工大学)
;
Department of Computer Science, University of Liverpool(计算机科学系,利物浦大学)
机构
*
Laboratory of Complex Systems Modeling and Simulation, School of Computer Science and Technology, Hangzhou Dianzi University(电子科技大学复杂系统建模与仿真实验室,计算机科学与技术学院)
;
Zhejiang Key Laboratory of Space Information Sensing and Transmission, Hangzhou Dianzi University(浙江省空间信息感知与传输重点实验室,电子科技大学)
;
Department of Psychological and Cognitive Sciences, Tsinghua University(清华大学心理与认知科学系)
MM-TS: Multi-Modal Temperature and Margin Schedules for Contrastive Learning with Long-Tail Data
MM-TS: 多模态温度和边距调度用于长尾数据的对比学习
Siarhei Sheludzko, Dhimitrios Duka, Bernt Schiele, Hilde Kuehne, Anna Kukleva
机构
*
University of Bonn(波恩大学)
;
MPI for Informatics, SIC(信息研究所)
;
Tuebingen AI Center/University of Tuebingen(图宾根人工智能中心/图宾根大学)
;
MIT-IBM Watson AI Lab(麻省理工-IBM Watson人工智能实验室)