Kestrel: Grounding Self-Refinement for LVLM Hallucination Mitigation
Kestrel: 为降低LVLM幻觉而引入自反思
机构 * UC Santa Cruz(加州大学圣克ruz分校) ; UC Berkeley(加州大学伯克利分校) ; UNC-Chapel Hill(北卡罗来纳大学教堂山分校) ; Apple(苹果公司)
专题命中 多模态Agent :multimodal(abstract);分类 cs.CV、cs.AI
AI总结 Kestrel提出一种无需训练的框架,通过显式视觉 grounding 与证据验证自反思机制减少LVLM幻觉,实验显示在POPE和MME-Hallucination基准上性能提升,同时提供透明的验证轨迹。
Comments 16 pages, 11 figures, 5 tables