A Systematic Evaluation of Positional Bias in Multi-Video Summarization with MLLMs
多视频摘要中位置偏差的系统评估:基于多模态大语言模型
机构 * School of Artificial Intelligence, Jilin University(吉林大学人工智能学院) ; Engineering Research Center of Knowledge-Driven Human-Machine Intelligence, Jilin University(知识驱动人机智能工程研究中心) ; International Center of Future Science, Jilin University(未来科学国际中心)
专题命中 评测与基准 :large language model(abstract);language model(abstract);分类 cs.CL
AI总结 本研究系统评估了多模态大语言模型在多视频摘要任务中的位置偏差,通过构建基准和三种互补指标揭示了领域与模型依赖的偏差特性,并分析了提示级缓解方法。