Diverse via bounded Agreement: Geometric Regularization for Multimodal Fusion
通过有界一致性实现多样性:多模态融合的几何正则化
Zixuan Xia, Hao Wang, Pengcheng Weng, Yanyu Qian, Yangxin Xu, William Dan, Fei Wang
机构
*
Department of Informatics University of Bern(伯尔尼大学信息学院)
;
College of Computing and Data Science Nanyang Technological University(南洋理工大学计算机与数据科学学院)
;
School of Software Engineering Xi’an Jiaotong University(西安交通大学软件工程学院)
机构
*
School of Artificial Intelligence, Nanjing University, China(南京大学人工智能学院)
;
National Key Laboratory for Novel Software Technology, Nanjing University, China(南京大学新型软件技术国家重点实验室)
机构
*
Qwen Large Model Application Team, Alibaba(阿里云大模型应用团队)
;
Alibaba University of Waterloo(阿里大学水力学院)
;
Vector Institute(向量研究所)
;
Zhejiang University(浙江大学)
MultiPUFFIN: A Multimodal Domain-Constrained Foundation Model for Molecular Property Prediction of Small Molecules
MultiPUFFIN:用于小分子性质预测的多模态领域约束基础模型
Idelfonso B. R. Nogueira, Carine M. Rebello, Mumin Enis Leblebici, Erick Giovani Sperandio Nascimento
机构
*
Department of Chemical Engineering, Norwegian University of Science and Technology (NTNU)(挪威科学与技术大学化学工程系)
;
Faculty of Industrial Engineering, KU Leuven(鲁文大学工业工程学院)
;
University of Surrey(萨里大学)
专题命中
多模态训练与对齐
:multimodal(title,abstract);cross-modal(abstract);multimodal foundation model(abstract);分类 cs.AI
PID-Guided Partial Alignment for Multimodal Decentralized Federated Learning
PID引导的多模态去中心化联邦学习部分对齐
Yanhang Shi, Xiaoyu Wang, Houwei Cao, Jian Li, Yong Liu
机构
*
Department of Electrical and Computer Engineering, Stony Brook University(石溪大学电气与计算机工程系)
;
Department of Applied Mathematics and Statistics and the Department of Computer Science, Stony Brook University(石溪大学应用数学与统计系和计算机科学系)
;
Department of Electrical and Computer Engineering, New York University(纽约大学电气与计算机工程系)
;
Department of Computer Science, New York Institute of Technology(纽约理工学院计算机科学系)
CommentsExperiments are inconclusive: The claim that architectures such as Chameleon or Emu would exhibit stronger gradient conflict is not supported by experiments or analysis, and all experiments are conducted on Janus-Pro without evaluation on other unified multimodal architectures
A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot Segmentation, Classification, and Deblurring
一种用于光片荧光显微镜的多模态3D基础模型实现少样本分割、分类和去模糊
Adina Scheinfeld, Haotan Zhang, Shang Mu, Rudolf L. M. van Herten, Lucas Stoffl, Ali Erturk, Zhuhao Wu, Johannes C. Paetzold
机构
*
Tri-Institutional Program in Computational Biology \& Medicine, Weill Cornell Medicine, New York, NY, USA Department of Radiology, Weill Cornell Medicine, New York, NY, USA Helen
;
Robert Appel Alzheimers Disease Research Institute, Feil Family Brain
;
Mind Research Institute, Weill Cornell Medicine, New York, NY, USA Graduate Program in Physiology, Biophysics
;
Systems Biology, Weill Cornell Medicine, New York, NY, USA Cornell Tech, New York, NY, USA Institute for Intelligent Biotechnologies (iBIO), Helmholtz Center Munich, Neuherberg, Germany Institute for Stroke
;
Dementia Research, Klinikum der Universität München, Ludwig-Maximilians University Munich, Munich, Germany
机构
*
Taizhou Institute of Science and Technology, Nanjing University of Science and Technology(泰州科技学院、南京理工大学)
;
Department of Intelligence Science, Xi’an Jiaotong-Liverpool University(智能科学系,西安交通大学利物浦大学)
;
School of Computer Science and Technology, Soochow University(计算机科学与技术学院,苏州大学)
;
Department of Statistical Sciences, University of Toronto(统计科学系,多伦多大学)
Semantics-Guided Multimodal Masked Autoencoder Pretraining for 3D BEV Object Detection
语义引导的多模态掩码自编码器预训练用于3D BEV目标检测
Prabuddhi Wariyapperuma, Rajitha de Silva, Marc Hanheide, Thomas Bohné, Leonardo Guevara
机构
*
University of Lincoln, Lincoln Centre for Autonomous Systems(林肯大学,林肯自主系统中心)
;
University of Cambridge, Institute for Manufacturing, Department of Engineering(剑桥大学,制造研究所,工程系)
机构
*
Key Laboratory of Child Development and Learning Science (Ministry of Education), School of Biological Sciences and Medical Engineering, Southeast University(儿童发展与学习科学重点实验室(教育部)、生物科学与医学工程学院、东南大学)
;
Department of Artificial Intelligence, Westlake University(人工智能学院、西湖大学)
;
Department of Artificial Intelligence, Vrije Universiteit Amsterdam(人工智能学院、阿姆斯特丹自由大学)
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
Baidu Inc(百度公司)
A Multimodal Framework for Dementia Detection via Linguistic and Acoustic Representation Learning
基于语言和声学表征学习的多模态痴呆检测框架
Loukas Ilias, Dimitris Askounis
机构
*
Decision Support Systems Laboratory, School of Electrical and Computer Engineering, National Technical University of Athens(决策支持系统实验室,电气与计算机工程学院,国家技术大学雅典)
机构
*
Department of Data Science \& AI, Faculty of Information Technology, Monash University, Melbourne, VIC 3800, Australia Alfred Health Radiology, Alfred Health, Melbourne, VIC 3004, Australia School of Translational Medicine, Faculty of Medicine, Nursing
;
Health Sciences, Monash University, Melbourne, VIC 3800, Australia Hong Kong Polytechnic University, Hong Kong SAR, China
CommentsEarly accepted by MICCAI 2026. This version of the contribution has been accepted for publication, after peer review (when applicable) but is not the Version of Record and does not reflect post-acceptance improvements, or any corrections
Image-Conditioned Instance Prompt Network for Referring Remote Sensing Image Segmentation
图像条件实例提示网络用于遥感图像指代分割
Biaoyu Ren, Qingsheng Wang, Cun Xu, Dingkang Yang, Wenxuan Wang
机构
*
School of Computer Science, Northwestern Polytechnical University, Xi'an, China(西北工业大学计算机科学学院,西安,中国)
;
College of Intelligent Robotics and Advanced Manufacturing, Fudan University, Shanghai, China(复旦大学智能机器人与先进制造学院,上海,中国)
;
Shenzhen Research Institute of Northwestern Polytechnical University, Shenzhen, China(西北工业大学深圳研究院,深圳,中国)
IVR-R1: Refining Trajectories through Iterative Visual-Grounded Reasoning in Reinforcement Learning
IVR-R1:通过强化学习中的迭代视觉基础推理优化轨迹
Chenghao Li, Fusheng Hao, Xikai Zhang, Likang Xiao, Yanwei Ren, Fuxiang Wu, Quan Chen, Liu Liu
机构
*
Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院)
;
School of Artificial Intelligence, Beihang University(北京航空航天大学人工智能学院)
;
Kuaishou Technology(快手科技)
;
Shenzhen Institute of Advanced Integration Technology, Shenzhen(深圳先进集成技术研究院)
机构
*
University of Macau(澳门大学)
;
Guangdong Institute of Intelligence Science and Technology(广东智能科学与技术研究院)
;
Peking University(北京大学)
;
Independent Researcher(独立研究员)
;
Institute of Science Tokyo(东京科学研究院)
;
Morgan Stanley(摩根大通)
;
Halmstad University(哈马碧大学)