Progressive Video Condensation with MLLM Agent for Long-form Video Understanding
逐步视频压缩与MLLM代理用于长视频理解
机构 * Zhejiang Key Laboratory of Space Information Sensing and Transmission, Hangzhou Dianzi University(浙江省空间信息感知与传输重点实验室,杭州电子科技大学)
专题命中 视频理解 :video understanding(title,abstract);long video(abstract);分类 cs.CV
AI总结 本文提出ProVCA,通过逐步缩小范围从粗略片段到精细帧,利用MLLM进行高效视频理解,实现高准确率。
Comments Accepted to ICME 2026