机构
*
Indian Institute of Technology Ropar(印度理工学院罗帕尔分校)
;
RoentGen Health(伦琴健康公司)
;
Deakin University(迪肯大学)
;
University of Central Florida(中佛罗里达大学)
;
Queensland University of Technology(昆士兰科技大学)
;
Murdoch University(莫道克大学)
机构
*
Department of Gynaecology and Obstetrics, The Affiliated Jiangyin Hospital of Nantong University(南通大学附属江阴医院妇产科)
;
Department of Oncology, the Affiliated Jiangyin Hospital of Nantong University(南通大学附属江阴医院肿瘤科)
;
Thrust of AI, Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)人工智能推力实验室)
;
FertiTech AI(生殖科技人工智能公司)
;
Department of Oncology, Suzhou Xiangcheng People’s Hospital(苏州市相城区人民医院肿瘤科)
;
Department of Biological Sciences and Bioinformatics, School of Science, Xi’an Jiaotong-Liverpool University(西交利物浦大学理学院生物科学与生物信息学系)
;
School of Artificial Intelligence and Computer Science, Nantong University(南通大学人工智能与计算机科学学院)
机构
*
Huazhong University of Science and Technology(华中科技大学)
;
Imperial Global Singapore, Imperial College London(帝国理工学院新加坡全球中心)
;
Nanyang Technological University(南洋理工大学)
;
National University of Singapore(新加坡国立大学)
;
Anhui University(安徽大学)
Revealing Training Data Exposure in Vision Language Large Models via Parameter Gradients
揭示视觉语言大模型中的训练数据暴露:基于参数梯度的方法
Zhihao Zhu, Hongyi Tang, Yi Yang, Ahmed Abbasi
机构
*
Department of Information Systems, Business Statistics and Operations Management (ISOM), Hong Kong University of Science and Technology, Hong Kong, China(信息系统、商业统计与运营管理系(ISOM),香港科技大学,香港,中国)
;
Department of IT, Analytics, and Operations, University of Notre Dame, Notre Dame, Indiana, USA(信息技术、分析与运营系,诺丁汉大学,诺丁汉,印第安纳州,美国)
An approach with Visual and Tabular Mamba to multimodal medical data using Mixed Fusion
一种基于视觉和表格Mamba的混合融合多模态医疗数据方法
Matheus B. Rocha, Gustavo B. Dettogni, Renato A. Krohling
机构
*
Labcin - Nature Inspired Computing Lab, Federal University of Esp\'irito Santo, Vit\'oria, Brazil PPGI - Graduate Program in Computer Science, Federal University of Esp\'irito Santo, Vit\'oria, Brazil
Confidence Calibration for Multimodal LLMs: An Empirical Study through Medical VQA
多模态大语言模型的置信度校准:基于医学视觉问答的实证研究
Yuetian Du, Yucheng Wang, Ming Kong, Tian Liang, Qiang Long, Bingdi Chen, Qiang Zhu
机构
*
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
School of Computer Science and Technology, Xidian University(西安电子科技大学计算机科学与技术学院)
;
Zhihui Medical Technology (Shanghai) Co., Ltd.(智汇医疗科技(上海)有限公司)
Hallucination Detection and Correction in Medical VLMs via Counter-Evidence Verification
基于反事实证据验证的医学视觉语言模型幻觉检测与纠正
Nan Zhou, Ke Zou, Meng Liu, Linchao He, Jiaqi Zhu, Yi Zhang, Hu Chen, Huazhu Fu
机构
*
College of Computer Science, Sichuan University(四川大学计算机科学学院)
;
Yong Loo Lin School of Medicine, National University of Singapore(新加坡国立大学杨潞龄医学院)
;
Key Laboratory of Data Protection and Intelligent Management, Ministry of Education, Sichuan University(四川大学数据保护与智能管理教育部重点实验室)
;
National Key Laboratory of Autonomous Intelligent Unmanned Systems, Beijing Institute of Technology(北京理工大学自主智能无人系统国家重点实验室)
;
Institute of High Performance Computing (IHPC), Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局高性能计算研究所)
Artemis: Anatomy-Resolved inTervention for Eliminating Multimodal NeuroImage confounderS
Artemis: 解剖分辨的干预方法用于消除多模态神经影像混杂因素
Siyuan Dai, Yang Du, Kun Zhao, Zhusuyi Chen, Heng Huang, Paul Thompson, Chao Shi, Haoteng Tang, Liang Zhan
机构
*
University of Pittsburgh(匹兹堡大学)
;
University of Maryland(马里兰大学)
;
University of Southern California(南加州大学)
;
Binghamton University(宾汉姆顿大学)
;
University of Texas Rio Grande Valley(德克萨斯大学里奥格兰德河谷分校)
机构
*
Department of Computer Science and Technology, Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)计算机科学与技术学院)
;
School of Intelligence Science and Engineering, College of Artificial Intelligence, Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)人工智能学院智能科学与工程学院)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
Guangdong Key Laboratory of Biomedical Measurements and Ultrasound Imaging, School of Biomedical Engineering, Shenzhen University Medical School, Shenzhen University(深圳大学医学部生物医学工程学院广东省生物医学测量与超声成像重点实验室)
;
Department of Radiology, The People’s Hospital of Guangxi Zhuang Autonomous Region, Guangxi Academy of Medical Sciences(广西壮族自治区人民医院放射科,广西医学科学院)
;
Shenzhen Sixth People’s Hospital (Nanshan Hospital), Huazhong University of Science and Technology Union Shenzhen Hospital(华中科技大学协和深圳医院(深圳市第六人民医院))
;
School of Basic Medical Sciences, Shenzhen University(深圳大学基础医学院)
;
Egypt-Japan University of Science and Technology (E-JUST)(埃及日本科技大学)
;
School of Biomedical Engineering, National-Regional Key Technology Engineering Laboratory for Medical Ultrasound, Guangdong Key Laboratory for Biomedical Measurements and Ultrasound Imaging, Shenzhen University Medical School(深圳大学医学部生物医学工程学院,国家地方联合医学超声关键技术工程实验室,广东省生物医学测量与超声成像重点实验室)
A report-grounded vision-language foundation model for colonoscopy from 280000 routine reports
基于28万份常规报告的肠镜报告驱动的视觉-语言基础模型
Jia Yu, Yan Zhu, Yili He, Zilong Wang, Xinyang Jiang, Peiyao Fu, Ruijie Yang, Tianyi Chen, Siyuan Li, Zhihua Wang, Fei Wu, Quanlin Li, Xian Yang, Pinghong Zhou, Shuo Wang
机构
*
Digital Medical Research Center, School of Basic Medical Sciences, Fudan University(复旦大学基础医学院数字医学研究中心)
;
Shanghai Collaborative Innovation Center of Endoscopy(上海内镜诊疗协同创新中心)
;
Zhejiang University(浙江大学)
;
Shanghai Institute for Advanced Study of Zhejiang University(浙江大学上海高等研究院)
;
Alliance Manchester Business School, The University of Manchester(曼彻斯特大学联盟曼彻斯特商学院)
;
Data Science Institute, Imperial College London(伦敦帝国理工学院数据科学研究所)
;
Microsoft Research Asia(微软亚洲研究院)
机构
*
Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系)
;
Institute of Medical Intelligence and XR, The Chinese University of Hong Kong(香港中文大学医学智能与扩展现实研究所)
Comments4 pages, 3 figures. Accepted at the 1st IJCAI Workshop on Safe Physical AI (SPAI 2026), held in conjunction with IJCAI-ECAI 2026, Bremen, Germany