Berna Kabadayi, Vanessa Sklyarova, Wojciech Zielonka, Justus Thies, Gerard Pons-Moll
机构
*
Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)
;
ETH Zürich(苏黎世联邦理工学院)
;
University of Tübingen(图宾根大学)
;
Technical University of Darmstadt(达姆施塔特工业大学)
;
Tübingen AI Center(图宾根人工智能中心)
;
Max Planck Institute for Informatics(马克斯·普朗克信息学研究所)
Exploring Motion-Language Alignment for Text-driven Motion Generation
探索文本与运动对齐以实现文本驱动的运动生成
Ruxi Gu, Zilei Wang, Wei Wang
机构
*
Department of Automation, University of Science and Technology of China(中国科学技术大学自动化系)
;
State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,北京通用人工智能研究院)
Multi-Modal Representation Learning via Semi-Supervised Rate Reduction for Generalized Category Discovery
通过半监督率减少实现多模态表示学习的通用类别发现
Wei He, Xianghan Meng, Zhiyuan Huang, Xianbiao Qi, Rong Xiao, Chun-Guang Li
机构
*
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
Intellifusion Inc., Shenzhen, P.R. China(深圳云天励飞技术股份有限公司)
专题命中
VLM训练与架构
:vision language model(abstract);分类 cs.CV
机构
*
Zachry Department of Civil
;
Environmental Engineering, Texas A\&M University
;
Department of Computer Science
;
Engineering, Texas A\&M University
GIFT: Global Irreplaceability Frame Targeting for Efficient Video Understanding
GIFT:全局不可替代性帧目标用于高效视频理解
Junpeng Ma, Sashuai Zhou, Guanghao Li, Xin Gao, Yue Cao, Hengyu Zeng, Yuxiang Yan, Zhibin Wang, Jun Song, Bo Zheng, Shanghang Zhang, Jian Pu
机构
*
Institute of Science and Technology for Brain-inspired Intelligence, Fudan University(复旦大学类脑智能科学与技术研究院)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(北京大学计算机学院多媒体信息处理国家重点实验室)
;
Zhejiang University(浙江大学)
;
Alibaba Group Holding Limited(阿里巴巴集团控股有限公司)
;
Future Living Lab of Alibaba(阿里巴巴未来生活实验室)
Cross-Modal Prototype Alignment and Mixing for Training-Free Few-Shot Classification
跨模态原型对齐与混合用于无训练少样本分类
Dipam Goswami, Simone Magistri, Gido M. van de Ven, Bartłomiej Twardowski, Andrew D. Bagdanov, Tinne Tuytelaars, Joost van de Weijer
机构
*
Department of Computer Science, Universitat Autònoma de Barcelona, Spain(巴塞罗那自治大学计算机科学系)
;
Computer Vision Center, Barcelona, Spain(巴塞罗那计算机视觉中心)
;
Media Integration and Communication Center, University of Florence, Italy(佛罗伦萨大学媒体整合与传播中心)
;
Bernoulli Institute, University of Groningen, the Netherlands(格罗宁根大学伯努利学院)
;
IDEAS Research Institute, Poland(波兰IDEAS研究所)
;
ESAT-PSI, KU Leuven, Belgium(比利时KU莱顿大学ESAT-PSI)