机构
*
School of Science and Engineering, The Chinese University of Hong Kong, Shenzhen(科学与工程学院,香港中文大学(深圳))
;
Shenzhen Future Network of Intelligence Institute(深圳未来网络智能研究院)
;
Wangxuan Institute of Computer Technology, Peking University(王轩计算机技术研究所,北京大学)
机构
*
Nanjing University(南京大学)
;
China Mobile Information Technology Co., Ltd.(中国移动信息科技有限公司)
;
Nanjing University of Information Science and Technology(南京信息科技大学)
Self-supervised Learning of Echocardiographic Video Representations via Online Cluster Distillation
通过在线聚类蒸馏实现心电图视频的自监督学习
Divyanshu Mishra, Mohammadreza Salehi, Pramit Saha, Olga Patey, Aris T. Papageorghiou, Yuki M. Asano, J. Alison Noble
机构
*
Department of Engineering Science, University of Oxford(工程科学系,牛津大学)
;
Nuffield Department of Women’s and Reproductive Health, University of Oxford(妇女与生殖健康系,牛津大学)
;
Fundamental AI Lab, University of Technology Nuremberg(基础人工智能实验室,纽伦堡技术大学)
;
University of Amsterdam(阿姆斯特丹大学)
Video Joint-Embedding Predictive Architectures for Facial Expression Recognition
视频联合嵌入预测架构用于面部表情识别
Lennart Eing, Cristina Luna-Jiménez, Silvan Mertes, Elisabeth André
机构
*
Chair for Human-Centered Artificial Intelligence(人中心人工智能教研室)
;
University of Augsburg(奥斯特拉赫大学)
;
Faculty of Computer Science(计算机科学学院)
;
Technical University of Applied Sciences Augsburg(应用科学大学奥斯特拉赫)
CommentsTo appear in 2025 Proceedings of the 13th International Conference on Affective Computing and Intelligent Interaction (ACII), submitted to IEEE. \c{opyright} 2025 IEEE
Pushing the Frontier of Audiovisual Perception with Large-Scale Multimodal Correspondence Learning
推动音频视觉感知前沿:大规模多模态对应学习
Apoorv Vyas, Heng-Jui Chang, Cheng-Fu Yang, Po-Yao Huang, Luya Gao, Julius Richter, Sanyuan Chen, Matt Le, Piotr Dollár, Christoph Feichtenhofer, Ann Lee, Wei-Ning Hsu
Comments15 pages, 6 figures, 1 table; accepted for AI-2025 Forty-fifth SGAI International Conference on Artificial Intelligence CAMBRIDGE, ENGLAND 16-18 DECEMBER 2025