CheXanatomy: Anatomy-Aware Vision-Language Modeling for Chest Radiographs
CheXanatomy: 面向胸部X光片的解剖感知视觉-语言建模
Sergios Gatidis, Curtis Langlotz, Christian Bluethgen
机构
*
Stanford Center for Artificial Intelligence in Medicine and Imaging, Stanford University(斯坦福大学医学与影像人工智能中心)
;
Department of Radiology, Stanford University(斯坦福大学放射学系)
Toward Generalizable Cognitive Impairment Detection with Speech-Based Multimodal Large Language Models
使用基于语音的多模态大语言模型实现可泛化的认知障碍检测
Yingchao Huang, Xin Wang, Yuhan Su, Shanshan Yao
机构
*
Faculty of Digital Innovation, Arts \& Sciences, Saskatchewan Polytechnic, Regina SK S4S 5X1, Canada
;
School of Basic Medical Sciences, Hebei University, Baoding 071000, China
;
Department of Civil \& Environmental Engineering
;
School of Mining \& Petroleum Engineering, University of Alberta, Edmonton AB T6G 2H5, Canada
EchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in Echocardiography
EchoSonar-R: 一种用于超声心动图疾病分类和报告生成的多视图推理增强模型
Darya Taratynova, Ahmed Aly, Numan Saeed, Mohammad Yaqub
机构
*
Division of Computing and Mathematical Sciences, Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI), Abu Dhabi, UAE(穆罕默德·本·扎耶德人工智能大学计算与数学科学系,阿布扎比,阿联酋)
Pathologist Attention-Aligned Report Generation for Prostate Histopathology
用于前列腺组织病理学的病理学家注意力对齐报告生成
Ruoyu Xue, Suryakant Singh, Souradeep Chakraborty, Pierre Marza, Oksana Yaskiv, Constantin Friedman, Natallia Sheuka, Paul Friedman, Bharat Ramlal, Beatrice Knudsen, Rajarsi Gupta, Joel Saltz, Prateek Prasanna, Gregory Zelinsky, Dimitris Samaras
机构
*
Department of Computer Science, Stony Brook University(纽约州立大学石溪分校计算机科学系)
;
Department of Biomedical Informatics, Stony Brook University(纽约州立大学石溪分校生物医学信息学系)
;
Université Paris-Saclay, CentraleSupélec, Gustave Roussy, INSERM, IHU PRISM, Cancer Data Science Unit(巴黎萨克雷大学、中央理工高等电力学院、古斯塔夫·鲁西研究所、法国国家健康与医学研究院、PRISM综合大学医院、癌症数据科学单元)
;
Université Paris-Saclay, CentraleSupélec, MICS Laboratory(巴黎萨克雷大学、中央理工高等电力学院、MICS实验室)
;
Department of Pathology and Laboratory Medicine, Northwell Health Laboratories(诺斯韦尔健康实验室病理与检验医学部)
;
Department of Pathology, University of Utah School of Medicine(犹他大学医学院病理系)
;
Department of Psychology, Stony Brook University(纽约州立大学石溪分校心理学系)
Comments11 pages, 4 figures, accepted for publication at the 29th International Conference on Medical Image Computing and Computer-Assisted Intervention (MICCAI 2026)
机构
*
Indian Institute of Technology Ropar(印度理工学院罗帕尔分校)
;
RoentGen Health(伦琴健康公司)
;
Deakin University(迪肯大学)
;
University of Central Florida(中佛罗里达大学)
;
Queensland University of Technology(昆士兰科技大学)
;
Murdoch University(莫道克大学)
机构
*
Department of Gynaecology and Obstetrics, The Affiliated Jiangyin Hospital of Nantong University(南通大学附属江阴医院妇产科)
;
Department of Oncology, the Affiliated Jiangyin Hospital of Nantong University(南通大学附属江阴医院肿瘤科)
;
Thrust of AI, Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)人工智能推力实验室)
;
FertiTech AI(生殖科技人工智能公司)
;
Department of Oncology, Suzhou Xiangcheng People’s Hospital(苏州市相城区人民医院肿瘤科)
;
Department of Biological Sciences and Bioinformatics, School of Science, Xi’an Jiaotong-Liverpool University(西交利物浦大学理学院生物科学与生物信息学系)
;
School of Artificial Intelligence and Computer Science, Nantong University(南通大学人工智能与计算机科学学院)
机构
*
Huazhong University of Science and Technology(华中科技大学)
;
Imperial Global Singapore, Imperial College London(帝国理工学院新加坡全球中心)
;
Nanyang Technological University(南洋理工大学)
;
National University of Singapore(新加坡国立大学)
;
Anhui University(安徽大学)
Revealing Training Data Exposure in Vision Language Large Models via Parameter Gradients
揭示视觉语言大模型中的训练数据暴露:基于参数梯度的方法
Zhihao Zhu, Hongyi Tang, Yi Yang, Ahmed Abbasi
机构
*
Department of Information Systems, Business Statistics and Operations Management (ISOM), Hong Kong University of Science and Technology, Hong Kong, China(信息系统、商业统计与运营管理系(ISOM),香港科技大学,香港,中国)
;
Department of IT, Analytics, and Operations, University of Notre Dame, Notre Dame, Indiana, USA(信息技术、分析与运营系,诺丁汉大学,诺丁汉,印第安纳州,美国)
An approach with Visual and Tabular Mamba to multimodal medical data using Mixed Fusion
一种基于视觉和表格Mamba的混合融合多模态医疗数据方法
Matheus B. Rocha, Gustavo B. Dettogni, Renato A. Krohling
机构
*
Labcin - Nature Inspired Computing Lab, Federal University of Esp\'irito Santo, Vit\'oria, Brazil PPGI - Graduate Program in Computer Science, Federal University of Esp\'irito Santo, Vit\'oria, Brazil
Confidence Calibration for Multimodal LLMs: An Empirical Study through Medical VQA
多模态大语言模型的置信度校准:基于医学视觉问答的实证研究
Yuetian Du, Yucheng Wang, Ming Kong, Tian Liang, Qiang Long, Bingdi Chen, Qiang Zhu
机构
*
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
School of Computer Science and Technology, Xidian University(西安电子科技大学计算机科学与技术学院)
;
Zhihui Medical Technology (Shanghai) Co., Ltd.(智汇医疗科技(上海)有限公司)
Hallucination Detection and Correction in Medical VLMs via Counter-Evidence Verification
基于反事实证据验证的医学视觉语言模型幻觉检测与纠正
Nan Zhou, Ke Zou, Meng Liu, Linchao He, Jiaqi Zhu, Yi Zhang, Hu Chen, Huazhu Fu
机构
*
College of Computer Science, Sichuan University(四川大学计算机科学学院)
;
Yong Loo Lin School of Medicine, National University of Singapore(新加坡国立大学杨潞龄医学院)
;
Key Laboratory of Data Protection and Intelligent Management, Ministry of Education, Sichuan University(四川大学数据保护与智能管理教育部重点实验室)
;
National Key Laboratory of Autonomous Intelligent Unmanned Systems, Beijing Institute of Technology(北京理工大学自主智能无人系统国家重点实验室)
;
Institute of High Performance Computing (IHPC), Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局高性能计算研究所)