Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark
看到场景才重要:通过场景感知的长视频基准揭示视频理解模型中的遗忘现象
机构 * CUHK (SZ)(香港中文大学(深圳)) ; University of Cambridge(剑桥大学) ; UESTC(电子科技大学) ; CUHK(香港中文大学) ; Shanghai Jiao Tong University(上海交通大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
AI总结 本文提出SceneBench基准,揭示视频理解模型在长场景上下文中的遗忘问题,并提出Scene-RAG方法提升性能2.50%。
Comments Accepted to CVPR 2026 (Highlight)
Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026