MEDLAYXPLAIN: Benchmarking the Expert-Lay Gap in Medical Vision-Language Models
MEDLAYXPLAIN: 医学视觉语言模型中的专家-外行差距基准测试
Han Jang, Junhyeok Lee, Songsoo Kim, Chae Young Lim, Hyeonjin Goh, Heeseong Eum, Kyu Sung Choi
机构
*
Seoul National University(首尔大学)
;
Seoul National University Hospital(首尔大学医院)
;
Seoul National University College of Medicine(首尔大学医学院)
;
Sungkyunkwan University School of Medicine(成均馆大学医学院)
Vision-language models for chest radiography do not always need the image
胸部X光片的视觉-语言模型并不总是需要图像
Mahshad Lotfinia, Sebastian Ziegelmayer, Lisa Adams, Daniel Truhn, Andreas Maier, Soroosh Tayebi Arasteh
机构
*
Pattern Recognition Lab, Friedrich-Alexander-Universität Erlangen-Nürnberg(弗里德里希-亚历山大-埃尔朗根-纽伦堡大学模式识别实验室)
;
Department of Diagnostic and Interventional Radiology, TUM University Clinic, School of Medicine and Health, Klinikum rechts der Isar, Technical University of Munich(慕尼黑工业大学医学院与健康学院伊萨尔河右岸医院诊断与介入放射学系)
;
Lab for AI in Medicine, RWTH Aachen University(亚琛工业大学医学人工智能实验室)
;
Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen(亚琛工业大学医院诊断与介入放射学系)
Scalable Training of Spatially Grounded 2D Vision-Language Models for Radiology
面向放射学的空间定位2D视觉-语言模型的可扩展训练
Yusuf Salcan, Simon Ging, Robin Tibor Schirrmeister, Philipp Arnold, Elmar Kotter, Behzad Bozorgtabar, Thomas Brox
机构
*
Computer Vision Group, University of Freiburg, Germany(德国弗莱堡大学计算机视觉组)
;
Department of Radiology, Medical Center -- University of Freiburg, Germany(德国弗莱堡大学医学中心放射科)
;
CRIION-AI Lab, Freiburg, Germany(德国弗莱堡CRIION-AI实验室)
GTA-Net: Cooperative Game Theory for Vision-Language Alignment in Chest X-Ray Report Generation
GTA-Net:合作博弈论在胸部X光报告生成中的视觉-语言对齐
Saif ur Rehman Khan, Imad Ahmed Waqar, Sebastian Vollmer, Andreas Dengel, Muhammad Nabeel Asim
机构
*
Department of Computer Science, Rhineland-Palatinate Technical University of Kaiserslautern-Landau(莱茵兰-普法尔茨凯泽斯劳滕-兰道工业大学计算机科学系)
;
German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)
;
IntelligentX GmbH
Attention-Spectrum Regularization for Replay-Free Continual Multimodal LLMs
注意力谱正则化:无回放的连续多模态大语言模型
Chuangxin Zhao, Canran Xiao, Siyuan Ma, Mengyao Lyu, Yanbiao Ma, Jun Xia, Guiguang Ding, Yang Liu
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Sun Yat-sen University(中山大学)
;
Nanyang Technological University(南洋理工大学)
;
Renmin University of China(中国人民大学)
;
Tsinghua University(清华大学)
机构
*
Kahlert School of Computing, University of Utah(犹他大学卡勒特计算学院)
;
Scientific Computing and Imaging Institute, University of Utah(犹他大学科学计算与成像研究所)
;
Department of Electrical and Computer Engineering, University of Utah(犹他大学电气与计算机工程系)
CommentsThis paper has been accepted for presentation at INTERSPEECH 2026 and as non-archival paper at ICML 2026 Workshop on Machine Learning for Audio
机构
*
Massachusetts Institute of Technology, USA(麻省理工学院)
;
National Taiwan University, Taipei, Taiwan(国立台湾大学)
;
National Taiwan University Artificial Intelligence Center of Research Excellence, Taipei, Taiwan(国立台湾大学人工智能研究中心)
;
Academia Sinica, Taiwan(台湾“中央”研究院)
;
National Yang Ming Chiao Tung University, Taiwan(阳明交通大学)
;
Signal Analysis and Interpretation Laboratory (SAIL), University of Southern California, USA(信号分析与解释实验室(SAIL),南加州大学)