A Patch-based Cross-view Regularized Framework for Backdoor Defense in Multimodal Large Language Models
基于补丁的跨视图正则化框架用于多模态大语言模型中的后门防御
Tianmeng Fang, Yong Wang, Zetai Kong, Zengzhen Su, Jun Wang, Chengjin Yu, Wei Wang
机构
*
Singapore Management University(新加坡管理大学)
;
China University of Mining and Technology(中国矿业大学)
;
The University of Melbourne(墨尔本大学)
;
Anhui University(安徽大学)
;
Xi’an Jiaotong University(西安交通大学)
;
Sun Yat-sen University(中山大学)
EI: Early Intervention for Multimodal Imaging based Disease Recognition
EI:基于多模态影像的疾病识别早期干预
Qijie Wei, Hailan Lin, Xirong Li
机构
*
Renmin University of China(中国人民大学)
;
Beijing Key Laboratory for Intelligent Diagnosis of Fundus Diseases and Drug-Device R&D and Translation(眼底疾病智能诊断与药械研发转化北京市重点实验室)
Chushan Zhang, Ruihan Lu, Jinguang Tong, Yikai Wang, Hongdong Li
机构
*
School of Computing, Australian National University(澳大利亚国立大学计算学院)
;
School of EECS, The University of Queensland(昆士兰大学电气工程与计算机科学学院)
;
FreiNexus
;
School of Artificial Intelligence, Beijing Normal University(北京师范大学人工智能学院)
A Multimodal Foundation Model of Spatial Transcriptomics and Histology for Biological Discovery and Clinical Prediction
一种结合空间转录组学和组织学的多模态基础模型用于生物发现和临床预测
Jinxi Xiang, Siyu Hou, Yuchen Li, Ryan Quinton, Xiaoming Zhang, Feyisope Eweje, Xiangde Luo, Yijiang Chen, Zhe Li, Colin Bergstrom, Ted Kim, Sierra Willens, Francesca Maria Olguin, Matthew Abikenari, Andrew Heider, Sanjeeth Rajaram, Joel Neal, Maximilian Diehn, Xiang Zhou, Ruijiang Li
机构
*
Department of Radiation Oncology, Stanford University School of Medicine(斯坦福大学医学院放射肿瘤学系)
;
Department of Statistics and Data Science, Yale University(耶鲁大学统计与数据科学系)
;
Department of Medicine (Oncology), Stanford University School of Medicine(斯坦福大学医学院医学系(肿瘤学))
;
Department of Pathology, Stanford University School of Medicine(斯坦福大学医学院病理学系)
;
Department of Neurosurgery, Stanford University School of Medicine(斯坦福大学医学院神经外科学系)
;
Stanford Institute for Human-Centered Artificial Intelligence(斯坦福大学以人为本人工智能研究所)
;
Perelman School of Medicine at the University of Pennsylvania(宾夕法尼亚大学佩雷尔曼医学院)
专题命中
其他多模态
:multimodal(title);multimodal foundation model(title);分类 cs.AI
Journal refProceedings of the 21st International Conference on Computer Vision Theory and Applications - Volume 3: VISAPP 2026; ISBN 978-989-758-804-4; ISSN 2184-4321, SciTePress, pages 353-364
ArchMap: Arch-Flattening and Knowledge-Guided Vision Language Model for Tooth Counting and Structured Dental Understanding
ArchMap:用于牙齿计数和结构牙科理解的拱形扁平化和知识引导的视觉语言模型
Bohan Zhang, Yiyi Miao, Taoyu Wu, Tong Chen, Ji Jiang, Zhuoxiao Li, Zhe Tang, Limin Yu, Jionglong Su
机构
*
Xi'an Jiaotong-Liverpool University(西交利物浦大学)
;
University of Liverpool(利物浦大学)
;
Zhejiang University of Technology(浙江工业大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))