Self-supervised Learning of Echocardiographic Video Representations via Online Cluster Distillation
通过在线聚类蒸馏实现心电图视频的自监督学习
Divyanshu Mishra, Mohammadreza Salehi, Pramit Saha, Olga Patey, Aris T. Papageorghiou, Yuki M. Asano, J. Alison Noble
机构
*
Department of Engineering Science, University of Oxford(工程科学系,牛津大学)
;
Nuffield Department of Women’s and Reproductive Health, University of Oxford(妇女与生殖健康系,牛津大学)
;
Fundamental AI Lab, University of Technology Nuremberg(基础人工智能实验室,纽伦堡技术大学)
;
University of Amsterdam(阿姆斯特丹大学)
Video Joint-Embedding Predictive Architectures for Facial Expression Recognition
视频联合嵌入预测架构用于面部表情识别
Lennart Eing, Cristina Luna-Jiménez, Silvan Mertes, Elisabeth André
机构
*
Chair for Human-Centered Artificial Intelligence(人中心人工智能教研室)
;
University of Augsburg(奥斯特拉赫大学)
;
Faculty of Computer Science(计算机科学学院)
;
Technical University of Applied Sciences Augsburg(应用科学大学奥斯特拉赫)
CommentsTo appear in 2025 Proceedings of the 13th International Conference on Affective Computing and Intelligent Interaction (ACII), submitted to IEEE. \c{opyright} 2025 IEEE
Pushing the Frontier of Audiovisual Perception with Large-Scale Multimodal Correspondence Learning
推动音频视觉感知前沿:大规模多模态对应学习
Apoorv Vyas, Heng-Jui Chang, Cheng-Fu Yang, Po-Yao Huang, Luya Gao, Julius Richter, Sanyuan Chen, Matt Le, Piotr Dollár, Christoph Feichtenhofer, Ann Lee, Wei-Ning Hsu
Comments15 pages, 6 figures, 1 table; accepted for AI-2025 Forty-fifth SGAI International Conference on Artificial Intelligence CAMBRIDGE, ENGLAND 16-18 DECEMBER 2025
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
Ye Liu, Zongyang Ma, Junfu Pu, Zhongang Qi, Yang Wu, Ying Shan, Chang Wen Chen
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
ARC Lab, Tencent PCG(腾讯PCG ARC实验室)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
vivo Mobile Communication Co.(vivo移动通信公司)
;
MindWingman Technology (Shenzhen) Co., Ltd.(深圳MindWingman技术有限公司)