To Agree or To Be Right? The Grounding-Sycophancy Tradeoff in Medical Vision-Language Models
同意还是正确?医学视觉-语言模型中的 grounding 与谄媚权衡
机构 * Department of Computer Science, The University of Texas at San Antonio(德克萨斯大学圣安东尼奥分校计算机科学系)
专题命中 幻觉与鲁棒性 :vision-language model(title,abstract);grounding(title,abstract);visual question answering(abstract);分类 cs.CV、cs.AI
AI总结 研究评估了六种医学视觉问答模型,揭示了 grounding 与谄媚之间的权衡关系,并提出三种新指标以评估模型的可靠性与安全性。