DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving
DRIVESPATIAL:自动驾驶中视觉语言模型时空智能的基准
机构 * University of Arkansas, USA(美国阿肯色大学) ; Google Research, Google(谷歌研究院) ; University of Liverpool, UK(英国利物浦大学) ; Max Planck Research School for Intelligent Systems(马克斯·普朗克智能系统研究学校)
专题命中 仿真评测 :autonomous driving(title,abstract);BEV(abstract,abstract_cn);分类 cs.CV
AI总结 提出DriveSpatial基准,通过多视角、时空推理任务评估视觉语言模型在自动驾驶中的场景构建、关系理解、时序推理和泛化能力,发现人类与模型间存在显著差距。