Unveiling and Bridging the Functional Perception Gap in MLLMs: Atomic Visual Alignment and Hierarchical Evaluation via PET-Bench
揭示和弥合MLLMs中的功能感知差距:通过PET-Bench实现原子视觉对齐和分层评估
机构 * School of Biomedical Engineering, Southern Medical University(生物医学工程学院,南方医科大学) ; School of Biomedical Engineering, Shanghai Jiaotong University(生物医学工程学院,上海交通大学) ; Department of Electronic Engineering, Chinese University of Hong Kong(电子工程系,中国香港大学) ; Faculty of Dentistry, The University of Hong Kong(牙科学院,香港大学) ; Department of Nuclear Medicine, The Second Affiliated Hospital of Guangzhou University of Chinese Medicine(核医学科,广州中医药大学第二附属医院) ; PET Center, Department of Nuclear Medicine, Guangdong Provincial People’s Hospital, Southern Medical University(PET中心,核医学科,广东省人民医院,南方医科大学) ; Department of Nuclear Medicine, Nanfang Hospital, Southern Medical University(核医学科,南芳医院,南方医科大学) ; Division of Nuclear Medicine and Molecular Imaging, Geneva University Hospitals(核医学与分子影像学部,日内瓦大学医院) ; Departments of Radiology, Physics, and Biomedical Engineering, The University of British Columbia(放射学、物理和生物医学工程系,不列颠哥伦比亚大学) ; Medical Artificial Intelligence Laboratory, Westlake University(医学人工智能实验室,西湖大学)
AI总结 本文提出AVA方法,通过原子视觉对齐解决MLLMs在功能成像中的感知差距,提升诊断准确性14.83%。
Comments 9 pages, 6 figures, 6 tables