HALP: Detecting Hallucinations in Vision-Language Models without Generating a Single Token
在不生成单个标记的情况下检测视觉-语言模型中的幻觉
机构 * Stony Brook University(石英溪大学) ; Toyota Technological Institute at Chicago(芝加哥丰田技术研究所)
专题命中 知识编辑与模型理解 :language model(title,abstract)
AI总结 通过探测模型内部表示在生成前检测视觉-语言模型的幻觉风险,展示不同架构中信息丰富的层和模态差异,并验证轻量级探测器在提升安全性和效率方面的潜力。
Journal ref The 19th Conference of the European Chapter of the Association for Computational Linguistics (EACL 2026)