Triple-Phase Multimodal Knowledge Aggregation Framework for Microbial Keratitis Subtype Diagnosis on Slit-Lamp Photography
基于裂隙灯摄影的微生物性角膜炎亚型诊断三相多模态知识聚合框架
Yiqing Wang, Maria A. Woodward, Ziyun Yang, N. Venkatesh Prajna, Chunming He, Leslie M. Niziol, Mercy Pawar, Ming-Chen Lu, Guillermo Amescua, Rachel Wozniak, Sejal Amin, Abinaya Krishnan, Prabhleen Kochar, Sina Farsiu
机构
*
Department of Biomedical Engineering, Duke University(杜克大学生物医学工程系)
;
Kellogg Eye Center, Department of Ophthalmology and Visual Sciences, University of Michigan(密歇根大学凯洛格眼科中心,眼科学与视觉科学系)
;
Department of Cornea and Refractive Surgery Services, Aravind Eye Care System(阿瓦因眼科医疗系统角膜与屈光手术部)
;
Bascom Palmer Eye Institute, Department of Ophthalmology, University of Miami Miller School of Medicine(迈阿密大学米勒医学院巴斯科姆·帕勒眼科研究所,眼科学系)
;
Flaum Eye Institute, Department of Ophthalmology, University of Rochester Medical Center(罗切斯特大学医学中心弗劳姆眼科研究所,眼科学系)
;
Department of Ophthalmology, Henry Ford Hospital(亨利福特医院眼科部)
;
Duke Eye Center, Duke University School of Medicine(杜克大学医学院杜克眼科中心)
Pluralis v0.1: Towards a Multicultural, Multimodal, Multilingual Benchmark for AI Risk and Reliability
Pluralis v0.1:迈向用于人工智能风险与可靠性的多元文化、多模态、多语言基准测试
Alicia Parrish, Rajat Shinde, Sanket Badhe, Xinyi Bai, Sree Bhargavi Balija, Hua-Rong Chu, Emilio Ferrara, Armstrong Foundjem, Rajat Ghosh, Aakash Gupta, Xuanli He, Ong Chen Hui, Minji Jung, Madhangi Karimanal, Faiza Khan Khattak, Boryoung Kim, Eugenia Kim, Liliya Lavitas, Seok Min Lim, Victor Lu, Jim Moirangthem, Dhivya Nagasubramanian, Deepak Pandita, Sita Rajagopal, Geetha Raju, Evgeniia Razumovskaia, Aravind Reddy, Federico Ricciuti, Nobin Sarwar, Sungpil Shin, Sunayana Sitaram, Snehal Thorat, Tharindu Cyril Weerasooriya, Jasmijn Bastings, Joachim Baumann, Kongtao Chen, Murali Emani, Mariya Hendriksen, Jiho Jin, Jun Seong Kim, Younghoon Ko, Alicja Kwasniewska, Minjae Lee, Tom Wei-cyuan Lin Kashyap Ramanandula Manjusha, Junho Myung, Junyeong Park, Roma Patel, Shyam Ratan, Sudarsun Santhiappan, Priyanka Suresh, Tuesday, Ksheeraj Sai Vepuri Laura Amortegui-Ordonez, Claire Dennis, Minsuk Kahng, Chris Knotz, Alice Oh, Balaraman Ravindran, Soojung Ryu William Bartholomew, Hiwot Tesfaye, Lora Aroyo
机构
*
Google DeepMind(谷歌DeepMind)
;
University of Alabama in Huntsville(阿拉巴马大学亨茨维尔分校)
;
Google(谷歌)
;
University of Missouri Columbia(密苏里大学哥伦比亚分校)
;
Chunghwa Telecom Laboratories(春木电信实验室)
;
University of Southern California(南加州大学)
;
Polytechnique Montreal(蒙特利尔理工学院)
;
Nutanix
;
ThinkEvolve Labs(ThinkEvolve实验室)
;
UCL(伦敦大学学院)
;
Infocomm Media Development Authority(信息通信媒体发展局)
;
Monark Health(Monark健康)
;
Seoul National University(首尔国立大学)
;
Microsoft(微软)
;
Centre for Responsible AI (CeRAI), Wadhwani School of Data Science and AI (WSAI), Indian Institute of Technology Madras(负责任人工智能中心(CeRAI)、瓦达威人工智能学校(WSAI)、印度理工学院马德拉斯分校)
;
University of Maryland, Baltimore County(马里兰大学巴尔的摩县分校)
;
Microsoft Research India(微软印度研究院)
;
Stanford University(斯坦福大学)
;
Argonne National Laboratory(阿贡国家实验室)
;
University of Oxford(牛津大学)
;
KAIST(韩国科学技术院)
;
Yonsei University(延世大学)
;
Amazon(亚马逊)
;
UIUC(伊利诺伊大学香槟分校)
;
Rochester Institute of Technology(罗切斯特理工学院)
;
Xenoscube Inc.(Xenoscube公司)
;
Korea AI Safety Institute (K-AISI)(韩国人工智能安全研究所(K-AISI))
;
MLCommons
;
CommonGround
;
Artifex Labs(Artifex实验室)
Framework and Multi-modal Dataset for Roadwork Zone Detection and Geo-localization
道路施工区域检测与地理定位的框架和多模态数据集
Zhiran Yan, Yutong Xin, S Shyam Shenoi, Rui Song, Gordon Elger
机构
*
Institute of Innovative Mobility (IIMo), Technical University Ingolstadt of Applied Sciences(应用科学英戈尔施塔特技术大学创新移动研究所)
;
Fraunhofer Institute for Transportation and Infrastructure Systems IVI(弗劳恩霍夫交通与基础设施系统研究所IVI)
;
Technical University of Munich(慕尼黑工业大学)
An Automated Multimodal Glaucoma Detection Framework Using ViT and a Stacking-Based Ensemble
使用ViT和基于堆叠的集成方法的自动多模态青光眼检测框架
Ishrat Jahan, Muhammad E. H Chowdhury, Murugappan Murugappan, Kanchon Kanti Podder, Tawsifur Rahman, Shrestha Datta, Md Sakib Abrar Hossain, Md Mosarrof Hossen, Yosra Magdi Salih Mekki, Sanjiban Sekhar Roy
机构
*
Department of Computer Science and Engineering, Shahjalal University of Science and Technology(肖哈尔大学科学与技术学院计算机科学与工程系)
;
Department of Electrical Engineering, Qatar University(卡塔尔大学电气工程系)
;
Department of Electronics and Communication Engineering, Kuwait College of Science and Technology(科威特科学与技术学院电子与通信工程系)
;
Department of Interdisciplinary Engineering, Kennesaw State University(肯尼斯州立大学跨学科工程系)
;
Department of Biomedical Engineering, School of Medicine, Johns Hopkins University(约翰霍普金斯大学医学院生物医学工程系)
;
Department of Biomedical Engineering, University of Oxford(牛津大学生物医学工程系)
;
Department of Computer Science and Engineering, Vellore Institute of Technology(维洛雷理工学院计算机科学与工程系)
RESOLVE: A Multi-Resolution and Multi-Modal Dataset for Roadside Cooperative Perception
RESOLVE:用于路边协同感知的多分辨率多模态数据集
Shaozu Ding, Linan Song, Marco De Vincenzi, Dajiang Suo
机构
*
The Polytechnic School, Arizona State University(亚利桑那州立大学理工学院)
;
Department of Computer Science and Engineering, New York University(纽约大学计算机科学与工程系)
Building a Multimodal Dataset of Academic Paper for Keyword Extraction
构建用于关键词提取的学术论文多模态数据集
Jingyu Zhang, Xinyi Yan, Yi Xiang, Yingyi Zhang, Chengzhi Zhang
机构
*
Department of Information Management, Nanjing University of Science and Technology(南京理工大学经济管理学院信息管理系)
;
Department of Archives and E-government, Soochow University(苏州大学社会学院档案与电子政务系)
Cross-view Multimodal Vision-Based Assessment Framework for Traditional Chinese Medicine Rehabilitation Training
跨视角多模态视觉评估框架用于中医康复训练
Francis Xiatian Zhang, Hao Yao, Shengxuan Chen, Hong Zhu, Hongxiao Jia, Sisi Zheng, Hubert P. H. Shum
机构
*
Department of Computer Science, Durham University(杜伦大学计算机科学系)
;
Institute for Regeneration and Repair, The University of Edinburgh(爱丁堡大学再生与修复研究所)
;
Ningbo Hospital of Traditional Chinese Medicine(宁波市中医院)
;
Department of Rehabilitation Medicine, The Gulou Hospital of Traditional Chinese Medicine(南京市鼓楼区中医院康复医学科)
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
QiYuanLab(启元实验室)
;
Tsinghua University(清华大学)
;
University of Electronic Science and Technology of China(电子科技大学)
机构
*
University of International Relations(国际关系学院)
;
Kedge Business School(凯致商学院)
;
Peking University(北京大学)
;
Tsinghua University(清华大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Capital Normal University(首都师范大学)
HEad and neCK TumOR (HECKTOR) 2025: Benchmark of Segmentation, Diagnosis, and Prognosis in Multimodal PET/CT
头颈肿瘤 (HECKTOR) 2025 挑战赛:多模态 PET/CT 中的分割、诊断与预后基准
Numan Saeed, Salma Hassan, Shahad Hardan, Lishan Cai, Xinglong Liang, Moona Mazher, Abdul Qayyum, Yansong Bu, Mengye Lyu, Yue Lin, Mingyuan Meng, Chuanyi Huang, Lisheng Wang, Dalal Chamseddine, Shamimeh Ahrari, Beining Wu, Yifei Chen, Fuyou Mao, Hao Zhang, Baixiang Zhao, Surajit Ray, Muzi Guo, Lei Xiang, Jakob Dexl, Michael Ingrisch, Adrien Depeursinge, Arman Rahmim, Mathieu Hatt, Vincent Andrearczyk, Mohammad Yaqub
机构
*
Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(穆罕默德·本·扎耶德人工智能大学)
;
Amsterdam UMC(阿姆斯特丹大学医学中心)
;
The Netherlands Cancer Institute(荷兰癌症研究所)
;
Radboud University Medical Centre(拉德堡德大学医学中心)
;
University College London(伦敦大学学院)
;
Imperial College London(帝国理工学院)
;
Shenzhen Technology University(深圳技术大学)
;
Shenzhen University(深圳大学)
;
Newland Digital Technology(新大陆数字技术)
;
The University of Sydney(悉尼大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
University Hospital, Nantes(南特大学医院)
;
Nantes Université, Centrale Nantes, CNRS, LS2N(南特大学、南特中央理工学院、法国国家科学研究中心、LS2N实验室)
;
Hangzhou Dianzi University(杭州电子科技大学)
;
Tsinghua University(清华大学)
;
Central South University(中南大学)
;
University of Glasgow(格拉斯哥大学)
;
China Mobile System Integration Co., Ltd.(中移系统集成有限公司)
;
Subtle Medical Inc.(Subtle Medical公司)
;
University Hospital, LMU Munich(慕尼黑大学医院)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
BC Cancer Research Institute(不列颠哥伦比亚癌症研究所)
;
HES-SO Valais-Wallis University of Applied Sciences and Arts(HES-SO瓦莱州应用科学与艺术大学)
;
Lausanne University Hospital (CHUV)(洛桑大学医院)
;
LaTIM, INSERM, UMR 1101, Univ Brest(LaTIM实验室、法国国家健康与医学研究院、UMR 1101、布雷斯特大学)
Comments17 pages, 4 figures, 4 tables. Overview paper for the HECKTOR 2025 challenge, held as a satellite event at MICCAI 2025. Challenge website: https://hecktor.grand-challenge.org/
MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning
MathVis-Fine:通过渐进式依赖引导训练将视觉监督与必要性对齐的多模态数学推理
Wanshi Xu, Haokun Zhao, Haidong Yuan, Songjun Cao, Long Ma
机构
*
School of ECE, Peking University(北京大学电子与计算机工程学院)
;
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与技术学院)
;
School of Software and Microelectronics, Peking University(北京大学软件与微电子学院)
;
Tencent Youtu Lab(腾讯优图实验室)
Comments8 pages, 6 figures. To appear in Proceedings of the 8th International Workshop on IoT Applications and Industry 5.0 (IoTI5 2026), co-located with IEEE DCOSS-IoT 2026, Reykjavik, Iceland, June 2026
Probing, Fusion, and Trustworthiness: A Systematic Evaluation of Foundation Model Representations for Multimodal Cancer Analysis
探测、融合与可信度:基础模型表示在多模态癌症分析中的系统评估
Jingyu Hu, Giuseppe Tripodi, Reed Naidoo, Sarah F. McGough, Tapabrata Chakraborti
机构
*
The Alan Turing Institute(艾伦·图灵研究所)
;
University of Bristol(布里斯托大学)
;
University of Manchester(曼彻斯特大学)
;
The Institute of Cancer Research(癌症研究所)
;
Genentech(基因泰克)
UrbanWell: Benchmarking Multimodal Large Language Models for Spatio-Temporal Urban Wellbeing Analytics
UrbanWell: 面向时空城市福祉分析的多模态大语言模型基准测试
Yanxin Xi, Xiang Su, Jie Feng, Yu Liu, Sasu Tarkoma, Pan Hui
机构
*
University of Helsinki(赫尔辛基大学)
;
Zhongguancun Academy(中关村学院)
;
University of Oxford(牛津大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))