机构
*
School of Electronic and Electrical Engineering, Shanghai University of Engineering Science(上海工程技术大学电子与电气工程学院)
;
Tencent Youtu Lab(腾讯云视频实验室)
;
ENT Institute and Department of Otorhinolaryngology, Eye & ENT Hospital of Fudan University(复旦大学耳鼻喉科医院耳鼻喉科研究所)
;
National University of Singapore(新加坡国立大学)
专题命中
多模态评测
:multimodal(title,abstract);multimodal foundation model(abstract);分类 cs.CV、cs.AI
机构
*
Kyoto University(京都大学)
;
NII LLMC(日本国立信息与通信技术研究所语言模型中心)
;
RIKEN AIP(日本理化学研究所先进理工研究所)
;
Case Western Reserve University(凯斯西储大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
The University of Osaka(大阪大学)
;
University of Tokyo(东京大学)
Recent Advances in Multi-modal 3D Intelligence: A Comprehensive Survey and Evaluation
多模态3D智能的最新进展:综合调查与评估
Yinjie Lei, Zixuan Wang, Feng Chen, Guoqing Wang, Peng Wang, Yang Yang
机构
*
College of Electronics and Information Engineering, Sichuan University(四川大学电子信息工程学院)
;
School of Computer Science, University of Adelaide(阿德莱德大学计算机科学学院)
;
School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院)
机构
*
Department of Computer Science and Technology, Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)计算机科学与技术系)
;
School of Intelligence Science and Engineering, College of Artificial Intelligence, Harbin Institute of Technology(哈尔滨工业大学智能科学与工程学院)
;
School of Biomedical Engineering, National-Regional Key Technology Engineering Laboratory for Medical Ultrasound, Guangdong Key Laboratory for Biomedical, Measurements and Ultrasound Imaging, Shenzhen University Medical School, Shenzhen University(深圳大学医学院生物医学工程学院、医学超声关键技术工程实验室、广东省生物医学测量与超声成像重点实验室)
Single-Channel Tissue Segmentation via Cross-Modal Distillation from Foundation Models
基于基础模型跨模态蒸馏的单通道组织分割
Sakib Mohammad, Jarin Ritu, Md Sakhawat Hossain
机构
*
Department of Engineering Technology(工程技术系)
;
Department of Electrical and Computer Engineering(电气与计算机工程系)
;
Department of Mechanical Engineering(机械工程系)
MM-Snowball: Evaluating and Mitigating Hallucination Snowballing in Multimodal Multi-Turn Dialogue
MM-Snowball:多模态多轮对话中的幻觉雪崩评估与缓解
Yue Jiang, Xue Jiang, Lihua Zhang, Zhiqiang Wang, Yuhang Lu, Peng Wang, Bo Han, Feng Zheng, Dingkang Yang
机构
*
College of Intelligent Robotics and Advanced Manufacturing, Fudan University(复旦大学智能机器人与先进制造学院)
;
Southern University of Science and Technology(南方科技大学)
;
TMLR Group, Hong Kong Baptist University(香港 Baptist 大学 TMLR 团体)
;
MM Lab, CUHK(CUHK 多模态实验室)
;
RAMS Lab, Huawei Technologies Co., Ltd.(华为技术有限公司 RAMS 实验室)
The Image Reconstruction Game: Drawing Common Ground Through Iterative Multimodal Dialogue
图像重建游戏:通过迭代多模态对话建立共同基础
Sherzod Hakimov, Mattia D'Agostini, Ivan Samodelkin, David Schlangen
机构
*
Computational Linguistics, Department of Linguistics University of Potsdam(波恩大学语言学系计算语言学部)
;
German Research Center for Artificial Intelligence (DFKI), Berlin(德国人工智能研究中心(DFKI)柏林)
A multimodal dataset of photoplethysmography and continuous behavioral responses to ASMR and nature videos
光电容积描记术和ASMR及自然视频连续行为反应的多模态数据集
Tushar Das, Daigo Hozaki, Koushlendra Kumar Singh, Hirohito M. Kondo
机构
*
Machine Vision & Intelligence Lab, National Institute of Technology Jamshedpur(机器视觉与智能实验室,jamshedpur国家理工学院)
;
School of Psychology, Chukyo University(心理学系,chukyo大学)
FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes
FigSIM:用于自杀迷因的细粒度自杀严重程度和比喻语言数据集
Liuliu Chen, Elise R. Carrotte, Brian E. Chapman, Jo Robinson, Mike Conway
机构
*
School of Computing and Information Systems, University of Melbourne, Australia(墨尔本大学计算与信息学院)
;
Orygen, The National Centre of Excellence in Youth Mental Health, Australia(奥里根青少年心理健康国家研究中心)
;
Centre for Youth Mental Health, University of Melbourne, Australia(墨尔本大学青少年心理健康中心)
;
O’Donnell School of Public Health, UT Southwestern Medical Center, United States(奥唐奈公共卫生学院,西南医学中心)
机构
*
National University of Singapore(国立新加坡大学)
;
Yunnan University(云南大学)
;
The Ohio State University(俄亥俄州立大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)