Where Does the Answer Come From? Benchmarking View-Level Visual Evidence Identification in Multi-View MLLMs for Autonomous Driving
答案从何而来?面向自动驾驶的多视角MLLMs中视角级视觉证据识别基准
机构 * University of Waterloo(滑铁卢大学)
专题命中 仿真评测 :autonomous driving(title,abstract);分类 cs.CV
AI总结 针对多视角自动驾驶场景,提出一个基准测试,评估多模态大模型在视觉问答中识别支持性相机视角的能力,包含122个冲突中心问题对,并区分视角选择与答案正确性。