Defending LVLMs Against Vision Attacks through Partial-Perception Supervision
Qi Zhou, Tianlin Li, Qing Guo, Dongxia Wang, Yun Lin, Yang Liu, Jin Song Dong
机构
*
College of Control Science and Engineering, Zhejiang University, China(控制科学与工程学院,浙江大学,中国)
;
Huzhou Institute of Industrial Control Technology, China(湖州工业控制技术研究所,中国)
;
School of Computer Science and Engineering, Nanyang Technological University, Singapore(计算机科学与工程学院,南洋理工大学,新加坡)
;
School of Computer Science, Shanghai Jiao Tong University, China(计算机科学学院,上海交通大学,中国)
;
School of Computing, National University of Singapore, Singapore(计算学院,新加坡国立大学,新加坡)
;
Research, Singapore(研究,新加坡)
专题命中
其他VLM
:vision language model(abstract);分类 cs.CV、cs.AI
iLearnRobot: An Interactive Learning-Based Multi-Modal Robot with Continuous Improvement
Kohou Wang, ZhaoXiang Liu, Lin Bai, Kun Fan, Xiang Liu, Huan Hu, Kai Wang, Shiguo Lian
机构
*
Unicom Data Intelligence(中国联通数据智能研究所)
;
Data Science & Artificial Intelligence Research Institute(数据科学与人工智能研究院)
;
China United Network Communications Group Corporation Limited(中国联合网络通信集团有限公司)
机构
*
School of Artificial Intelligence, Nanjing University(人工智能学院,南京大学)
;
National Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家实验室,南京大学)
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Hamad Bin Khalifa University(哈马德·本·卡西姆大学)
;
Central South University(中南大学)
Bound by semanticity: universal laws governing the generalization-identification tradeoff
Marco Nurisso, Jesseba Fernando, Raj Deshpande, Alan Perotti, Raja Marjieh, Steven M. Frankland, Richard L. Lewis, Taylor W. Webb, Declan Campbell, Francesco Vaccarino, Jonathan D. Cohen, Giovanni Petri
机构
*
Dipartimento di Scienze Matematiche, Politecnico di Torino(都灵理工大学数学科学系)
;
CENTAI Institute(CENTAI研究院)
;
Network Science Institute, Northeastern University(东北大学网络科学研究所)
;
Institute for Experiential AI, Northeastern University(东北大学体验人工智能研究所)
;
NP Lab, Network Science Institute, Northeastern University London(东北大学伦敦网络科学研究所NP实验室)
;
Department of Psychology, Princeton University(普林斯顿大学心理学系)
;
Program in Cognitive Science, Dartmouth College(达特茅斯学院认知科学项目)
;
Department of Psychology, University of Michigan(密歇根大学心理学系)
;
Microsoft Research(微软研究院)
;
Princeton Neuroscience Institute(普林斯顿神经科学研究所)
;
Department of Physics, Northeastern University(东北大学物理系)
PunchBench: Benchmarking MLLMs in Multimodal Punchline Comprehension
Kun Ouyang, Yuanxin Liu, Shicheng Li, Yi Liu, Hao Zhou, Fandong Meng, Jie Zhou, Xu Sun
机构
*
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机学院,北京大学)
;
WeChat AI, Tencent Inc., China(微信AI,腾讯公司,中国)
专题命中
其他VLM
:multimodal large language model(abstract);分类 cs.CV、cs.AI
CommentsThis is the camera-ready version for ACL 2025