Co-Learning for Missing Arbitrary Modalities in Multi-modal Classification
多模态分类中用于缺失任意模态的协同学习
Francisco Mena, Dino Ienco, Roberto Interdonato, Cassio F. Dantas, Simon Besnard
机构
*
GFZ Helmholtz Center for Geosciences(德国波茨坦地学研究中心亥姆霍兹中心)
;
INRAE, UMR TETIS, University of Montpellier(法国蒙彼利埃大学农业环境与生态研究院、蒙彼利埃大学TETIS联合研究单位)
;
INRIA, EVERGREEN, University of Montpellier(法国蒙彼利埃大学信息与自动化研究所、EVERGREEN实验室)
;
CIRAD, UMR TETIS, University of Montpellier(法国蒙彼利埃大学国际农业研究磋商组织、蒙彼利埃大学TETIS联合研究单位)
机构
*
Dept. of Electrical and Electronic Engineering, The Hong Kong Polytechnic University(香港理工大学电子与电气工程系)
;
Speech, Language, and Cognition Laboratory, The University of Hong Kong(香港大学语音、语言与认知实验室)
Large Language Models as Unified Multimodal Learners for Clinical Prediction
作为临床预测统一多模态学习者的大语言模型
Ajay Madhavan Ravichandran, Bilgin Osmandoja, Klemens Budde, Klaus Netter, Tobias Strapatsas, Aljoscha Burchardt, Sebastian Möller, Roland Roller
机构
*
German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心(DFKI))
;
Charité Universitätsmedizin Berlin(柏林夏里特大学医学中心)
;
DNC Information Management GmbH(DNC信息管理有限公司)
;
Klinik für Akut- und Notfallmedizin, Asklepios Klinikum Harburg(哈堡阿斯克勒庇俄斯医院急性与急诊医学科)
;
Technical University Berlin(柏林工业大学)
OmniFocus: Query-Guided Modality-Balanced Token Compression for Omni-Modal Large Language Models
OmniFocus:用于全模态大语言模型的查询引导模态平衡令牌压缩
Shijie Cao, Qingyu Zhang, Boxi Yu, Yuzhong Zhang, Boxi Cao, Yaojie Lu, Hongyu Lin, Xianpei Han, Le Sun
机构
*
School of Advanced Interdisciplinary Sciences, University of Chinese Academy of Sciences(中国科学院大学先进交叉科学学院)
;
Chinese Information Processing Laboratory, Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所中文信息处理实验室)
;
University of Limerick(利默里克大学)
;
CUHK, Shenzhen(香港中文大学(深圳))
Cross4D-JEPA: Dense Cross-modal Correspondence Distillation for 4D Point Cloud Representation Learning
Cross4D-JEPA: 用于4D点云表示学习的密集跨模态对应蒸馏
Trung Thanh Nguyen, Hai Nguyen-Truong, Tu Vo, Hoang M. Truong, Tuan-Anh Vu
机构
*
Nagoya University(名古屋大学)
;
Northeastern University(东北大学)
;
KC Machine Learning Lab(KC机器学习实验室)
;
University of Science, Vietnam National University Ho Chi Minh City(胡志明市越南国立大学理科大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
机构
*
Zhejiang University(浙江大学)
;
DAMO Academy, Alibaba Group(阿里巴巴集团达摩院)
;
Hupan Lab(华平实验室)
;
Huazhong University of Science and Technology(华中科技大学)
;
East China Normal University(华东师范大学)
;
Shanghai Jiao Tong University(上海交通大学)
An approach with Visual and Tabular Mamba to multimodal medical data using Mixed Fusion
一种基于视觉和表格Mamba的混合融合多模态医疗数据方法
Matheus B. Rocha, Gustavo B. Dettogni, Renato A. Krohling
机构
*
Labcin - Nature Inspired Computing Lab, Federal University of Esp\'irito Santo, Vit\'oria, Brazil PPGI - Graduate Program in Computer Science, Federal University of Esp\'irito Santo, Vit\'oria, Brazil
Plug-and-Adapt: Multimodal Coreference Resolution at First Sight with a Pretrained Alignment Model
即插即适应:基于预训练对齐模型的首眼多模态指代消解
Jinghan Wu, Jing Li, Ivor W. Tsang, Xuetao Zhang
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi'an Jiaotong University(西安交通大学人工智能与机器人研究所人机混合增强智能全国重点实验室)
;
Centre for Frontier AI Research and Institute of High-Performance Computing, Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局前沿人工智能研究中心与高性能计算研究所)
Information-Theoretic Decomposition for Multimodal Interaction Learning
多模态交互学习的信息论分解
Zequn Yang, Yake Wei, Haotian Ni, Zhihao Xu, Di Hu
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China, Beijing(中国人民大学人工智能学院,北京)
;
Beijing Key Laboratory of Research on Large Models(北京大模型研究关键实验室)
;
Engineering Research Center of Next-Generation Intelligent Search(下一代智能搜索与推荐工程研究中心)
;
Beihang University, Beijing(北航,北京)
;
Gaotu Techedu Inc.(高图科技有限公司)
机构
*
Department of Computer Science and Technology, Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)计算机科学与技术学院)
;
School of Intelligence Science and Engineering, College of Artificial Intelligence, Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)人工智能学院智能科学与工程学院)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
Guangdong Key Laboratory of Biomedical Measurements and Ultrasound Imaging, School of Biomedical Engineering, Shenzhen University Medical School, Shenzhen University(深圳大学医学部生物医学工程学院广东省生物医学测量与超声成像重点实验室)
;
Department of Radiology, The People’s Hospital of Guangxi Zhuang Autonomous Region, Guangxi Academy of Medical Sciences(广西壮族自治区人民医院放射科,广西医学科学院)
;
Shenzhen Sixth People’s Hospital (Nanshan Hospital), Huazhong University of Science and Technology Union Shenzhen Hospital(华中科技大学协和深圳医院(深圳市第六人民医院))
;
School of Basic Medical Sciences, Shenzhen University(深圳大学基础医学院)
;
Egypt-Japan University of Science and Technology (E-JUST)(埃及日本科技大学)
;
School of Biomedical Engineering, National-Regional Key Technology Engineering Laboratory for Medical Ultrasound, Guangdong Key Laboratory for Biomedical Measurements and Ultrasound Imaging, Shenzhen University Medical School(深圳大学医学部生物医学工程学院,国家地方联合医学超声关键技术工程实验室,广东省生物医学测量与超声成像重点实验室)
When Two Tracers Disagree: An Investigation of Multimodal Fusion for Clinical PET/CT Segmentation
当两种示踪剂意见不合时:临床PET/CT分割的多模态融合研究
Jack A. Johnson, Bartłomiej W. Papież
机构
*
University of Oxford(牛津大学)
;
Nuffield Department of Medicine(纳菲尔德医学系)
;
Department of Oncology(肿瘤学系)
;
Big Data Institute(大数据研究所)
;
Nuffield Department of Population Health(纳菲尔德人口健康系)
Comments10 pages (8 pages main content and 2 pages of references), 2 figures, 2 tables, accepted to MICCAI 2026 Cancer Prevention, Detection, and IntervenTion (CaPTion) Workshop
机构
*
Center for AI and Data Science (CAIDAS), Julius-Maximilians-Universität Würzburg(维尔茨堡大学人工智能与数据科学中心)
;
Institute for Computational Imaging and AI in Medicine (CompAI), Technical University of Munich(慕尼黑工业大学计算成像与医学人工智能研究所)
;
Pattern Recognition Lab, Friedrich-Alexander Universität Erlangen-Nürnberg(埃尔朗根-纽伦堡大学模式识别实验室)
;
Department Artificial Intelligence in Biomedical Engineering (AIBE), Friedrich-Alexander-Universität Erlangen-Nürnberg(埃尔朗根-纽伦堡大学生物医学工程人工智能系)
GeoUniPR: A Geometry-Consistent Unified Framework for Cross-Modal Place Recognition
GeoUniPR:用于跨模态地点识别的几何一致性统一框架
Wonbong Kim, Jiatong Xiao, Rui Li, Xufei Wang, Qiwen Gu, Junqiao Zhao, Chen Ye, Guang Chen
机构
*
School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院)
;
Shanghai Research Institute for Intelligent Autonomous System, Tongji University(同济大学上海智能自主系统研究院)
;
Shanghai Innovation Institute(上海创新研究院)