机构
*
School of Computing and Artificial Intelligence(计算与人工智能学院)
;
Engineering Research Center of Intelligent Finance Ministry of Education(教育部长智能金融工程研究中心)
;
Shenzhen Institutes of Advanced Technology Chinese Academy of Sciences(中国科学院深圳先进技术研究院)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Shanghai Academy of AI for Science(上海人工智能科学研究院)
;
Artificial Intelligence Innovation and Incubation Institute Fudan University(复旦大学人工智能创新与孵化院)
;
Artificial Intelligence and Digital Finance Key Laboratory of Sichuan Province(四川省人工智能与数字金融重点实验室)
Multi-modal Co-learning for Earth Observation: Enhancing single-modality models via modality collaboration
Francisco Mena, Dino Ienco, Cassio F. Dantas, Roberto Interdonato, Andreas Dengel
机构
*
Department of Computer Science, University of Kaiserslautern-Landau (RPTU)(科斯拉尔特伦大学计算机科学系)
;
SDS, German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心(DFKI))
;
INRAE, UMR TETIS, University of Montpellier(蒙彼利埃大学UMR TETIS)
;
CIRAD, UMR TETIS, University of Montpellier(蒙彼利埃大学UMR TETIS)
;
INRIA, EVERGREEN, University of Montpellier(蒙彼利埃大学)
MOTOR: Multimodal Optimal Transport via Grounded Retrieval in Medical Visual Question Answering
Mai A. Shaaban, Tausifa Jan Saleem, Vijay Ram Papineni, Mohammad Yaqub
机构
*
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
Department of Mathematics and Computer Science, Faculty of Science, Alexandria University(亚历山大大学数学与计算机科学系)
;
Sheikh Shakhbout Medical City(谢赫·沙赫布OUT医疗城)
CommentsIn Proceedings of the 2nd ACM Workshop in AI-powered Question and Answering Systems (AIQAM '25), October 27-28, 2025, Dublin, Ireland. ACM, New York, NY, USA, 8 pages. https://doi.org/10.1145/3746274.3760393
Knowledge-based Visual Question Answer with Multimodal Processing, Retrieval and Filtering
Yuyang Hong, Jiaqi Gu, Qi Yang, Lubin Fan, Yue Wu, Ying Wang, Kun Ding, Shiming Xiang, Jieping Ye
机构
*
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所MAIS)
;
Alibaba Cloud Computing(阿里巴巴云计算)
Robult: Leveraging Redundancy and Modality Specific Features for Robust Multimodal Learning
Duy A. Nguyen, Abhi Kamboj, Minh N. Do
机构
*
Siebel School of Computing and Data Science, UIUC, US(UIUC计算机与数据科学学院)
;
Department of Electrical and Computer Engineering, UIUC, US(UIUC电气与计算机工程学院)
;
VinUni-Illinois Smart Health Center, VinUniversity, Vietnam(Vin大学-伊利诺伊智能健康中心)
Street-Level Geolocalization Using Multimodal Large Language Models and Retrieval-Augmented Generation
Yunus Serhat Bicakci, Joseph Shingleton, Anahid Basiri
机构
*
Vocational School of Social Sciences, Marmara University(马尔马拉大学社会科学职业学校)
;
Geospatial Data Science Group, School of Geographical & Earth Sciences, University of Glasgow(格拉斯哥大学地理与地球科学学院空间数据科学小组)
机构
*
Department of Electrical and Computer Engineering, Isfahan University of Technology(电气与计算机工程系,伊斯法罕技术大学)
;
SDU Health Informatics and Technology, The Maersk Mc-Kinney Moller Institute, University of Southern Denmark(南部丹麦大学健康信息学与技术,马士基麦金尼莫勒研究所)
;
Geriatric Research Unit, Department of Clinical Research, University of Southern Denmark(老年医学研究单元,临床研究系,南部丹麦大学)