VLM-SlideEval: Evaluating VLMs on Structured Comprehension and Perturbation Sensitivity in PPT
Hyeonsu Kang, Emily Bao, Anjan Goswami
专题命中
图文多模态
:multimodal(abstract);分类 cs.CV、cs.AI
Comments39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: Evaluating the Evolving LLM Lifecycle - Benchmarks, Emergent Abilities, and Scaling
Frame-Difference Guided Dynamic Region Perception for CLIP Adaptation in Text-Video Retrieval
Jiaao Yu, Mingjie Han, Tao Gong, Jian Zhang, Man Lan
机构
*
School of Computer Science and Technology, East China Normal University, China(上海师范大学计算机科学与技术学院)
;
School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学与技术学院)
机构
*
School of Computer Science and Technology, University of Science and Technology of China(计算机科学与技术学院,中国科学技术大学)
;
State Key Lab of Processors, Institute of Computing Technology, Chinese Academy of Sciences(处理器国家重点实验室,中国科学院计算技术研究所)
;
Department of Computer Science, National University of Singapore(计算机科学系,新加坡国立大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Intelligent Software Research Center, Institute of Software, Chinese Academy of Sciences(软件智能研究中心,中国科学院软件研究所)
I Spy With My Model's Eye: Visual Search as a Behavioural Test for MLLMs
John Burden, Jonathan Prunty, Ben Slater, Matthieu Tehenan, Greg Davis, Lucy Cheke
机构
*
Leverhulme Centre for the Future of Intelligence, University of Cambridge(未来智能研究中心、剑桥大学)
;
Department of Engineering, University of Cambridge(工程系、剑桥大学)
;
Department of Psychology, University of Cambridge(心理学系、剑桥大学)
;
Department of Computer Science, University of Cambridge(计算机科学系、剑桥大学)
Learning to Detect Unknown Jailbreak Attacks in Large Vision-Language Models
Shuang Liang, Zhihao Xu, Jialing Tao, Hui Xue, Xiting Wang
专题命中
图文多模态
:multi-modal(abstract);分类 cs.CV、cs.AI
CommentsWithdrawn due to an accidental duplicate submission. This paper (arXiv:2510.15430) was unintentionally submitted as a new entry instead of a new version of our previous work (arXiv:2508.09201)
机构
*
Department of Computer Vision(计算机视觉系)
;
Mohamed bin Zayed University of Artificial Intelligence(马尔代夫比兹人工智能大学)
;
Department of Machine Learning(机器学习系)
;
Corniche Hospital, Abu Dhabi Health Services Company (SEHA)(阿布扎赫尔医院,阿布扎赫健康服务公司(SEHA))
MEGC2025: Micro-Expression Grand Challenge on Spot Then Recognize and Visual Question Answering
Xinqi Fan, Jingting Li, John See, Moi Hoon Yap, Wen-Huang Cheng, Xiaobai Li, Xiaopeng Hong, Su-Jing Wang, Adrian K. Davision
机构
*
Department of Computing and Mathematics, Manchester Metropolitan University(计算与数学系,曼彻斯特 Metropolitan 大学)
;
State Key Laboratory of Cognitive Science and Mental Health, Institute of Psychology, Chinese Academy of Sciences(认知科学与心理健康国家重点实验室,心理学研究所,中国科学院)
;
Department of Psychology, University of the Chinese Academy of Sciences(心理学系,中国科学院大学)
;
National Taiwan University(台湾大学)
;
Zhejiang University(浙江大学)
;
University of Oulu(奥卢大学)
;
Harbin Institute of Technology(哈尔滨工业大学)
专题命中
图文多模态
:multimodal(abstract);分类 cs.CV、cs.MM
CommentsMicro-Expression Grand Challenge (MEGC) at ACM MM 2025
机构
*
State Key Laboratory of Synthetical Automation for Process Industries, Northeastern University, Shenyang, China(合成过程工业综合自动化国家重点实验室,东北大学,沈阳,中国)
;
University of Surrey(Surrey大学)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)
;
School of Computer Science, The University of Adelaide(阿德莱德大学计算机学院)
;
College of Computing & Data Science, Nanyang Technological University(南洋理工大学计算机与数据科学学院)
;
Surrey Institute for People-Centred Artificial Intelligence, and Centre for Vision, Speech and Signal Processing, University of Surrey(Surrey人本人工智能研究所,以及视觉、语音和信号处理中心,Surrey大学)
专题命中
图文多模态
:multimodal(abstract);分类 cs.CV、cs.AI
Comments63 pages (main paper and supplementary material), 39 figures, 58 tables
MMLongBench: Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly
Zhaowei Wang, Wenhao Yu, Xiyu Ren, Jipeng Zhang, Yu Zhao, Rohit Saxena, Liang Cheng, Ginny Wong, Simon See, Pasquale Minervini, Yangqiu Song, Mark Steedman
机构
*
CSE Department, HKUST(香港科技大学计算机科学与工程系)
;
Tencent AI Seattle Lab(腾讯AI西雅图实验室)
;
University of Edinburgh(爱丁堡大学)
;
NVIDIA AI Technology Center (NVAITC), NVIDIA, Santa Clara, USA(英伟达圣克拉拉人工智能技术中心)