SAVER: Mitigating Hallucinations in Large Vision-Language Models via Style-Aware Visual Early Revision
SAVER:通过风格感知视觉早期修正减轻大型视觉语言模型中的幻觉
机构 * ROSE Lab, Interdisciplinary Graduate Programme, Nanyang Technological University, Singapore(南洋理工大学罗思实验室,跨学科研究生项目,新加坡) ; ROSE Lab, School of Electrical and Electronic Engineering, Nanyang Technological University, Singapore(南洋理工大学罗思实验室,电子与电气工程学院,新加坡) ; City University of Hong Kong, Hong Kong SAR(香港城市大学,香港特别行政区) ; Shanghai Jiao Tong University, China(上海交通大学,中国) ; Singapore University of Technology and Design, Singapore(新加坡科技设计大学,新加坡) ; VinUniversity, Hanoi, Vietnam(越南文大学,河内,越南)
AI总结 研究大型视觉语言模型幻觉问题,构建含照片及风格化图像数据集并对比,提出SAVER机制,利用早期层反馈基于视觉注意力模式动态调整输出,减轻风格化图像引起的幻觉,实验证明其性能先进。
Comments Accepted at AAAI 2026. 24 pages, 10 figures. Code: https://github.com/llizhaoxu/SAVER