机构
*
The Chinese University of Hong Kong(香港中文大学)
;
The Third Affiliated Hospital of Sun Yat-sen University(中山大学附属第三医院)
;
University of Cambridge(剑桥大学)
;
University of California, Davis(加州大学戴维斯分校)
Ye Yuan, Kehan Chen, Xinqiang Yu, Wentao Xu, Heng Wang, Libo Huang, Chuanguang Yang, Yan Huang, Jiawei He, Zhulin An
机构
*
School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所人工智能安全国家重点实验室)
;
National Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
XYZ Embodied AI(XYZ具身人工智能)
XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations
XR-1:通过学习统一的视觉-运动表示实现多功能的视觉-语言-动作模型
Shichao Fan, Kun Wu, Zhengping Che, Xinhua Wang, Di Wu, Fei Liao, Ning Liu, Yixue Zhang, Zhen Zhao, Zhiyuan Xu, Meng Li, Qingjie Liu, Shanghang Zhang, Min Wan, Jian Tang
机构
*
Beijing Innovation Center of Humanoid Robotics, Beijing, China(北京人形机器人创新中心,北京,中国)
;
School of Mechanical Engineering and Automation, Beihang University, Beijing, China(北京航空航天大学机械工程及自动化学院,北京,中国)
;
State Key Laboratory of Virtual Reality Technology and Systems, SCSE, Beihang University, Beijing, China(虚拟现实技术与系统国家重点实验室,SCSE,北京航空航天大学,北京,中国)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University, Beijing, China(多媒体信息处理国家重点实验室,计算机科学学院,北京大学,北京,中国)
CommentsWe have further refined the benchmark construction and experimental presentation to improve clarity and consistency. The revised version includes updated task design, food-resource data, and evaluation details to better align the benchmark with the intended food resource referral setting. These changes provide a more precise presentation of the experimental findings
机构
*
Pengcheng Laboratory(鹏城实验室)
;
School of Computer Science and Cyber Engineering(计算机科学与网络工程学院)
;
Guangzhou University(广州大学)
;
Southern University of Science and Technology(南方科技大学)
机构
*
Department of Civil and Environmental Engineering, University of Wisconsin-Madison(威斯康星大学麦迪逊分校土木与环境工程系)
;
Lyles School of Civil and Construction Engineering, Purdue University(普渡大学莱尔斯土木与建设工程学院)
;
Department of Civil Engineering, McGill University(麦吉尔大学土木工程系)
Are Neuro-Inspired Multi-Modal Vision-Language Models Resilient to Membership Inference Privacy Leakage?
受神经启发的多模态视觉-语言模型对成员推断隐私泄露是否具有弹性?
David Amebley, Sayanton Dibbo
机构
*
The University of Alabama(阿拉巴马大学)
;
Alabama Center for the Advancement of AI(阿拉巴马人工智能 advancement 中心)
;
Trustworthy AI Lab(可信人工智能实验室)
;
Department of Computer Science, The University of Alabama(计算机科学系)
机构
*
KTH Royal Institute of Technology(皇家理工学院)
;
Swiss Federal Institute of Technology Lausanne(洛桑联邦理工学院)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院)
;
CISPA Helmholtz Center for Information Security(信息安全赫尔姆霍兹中心)
;
RISE Research Institutes of Sweden(瑞典RISE研究机构)
;
Halmstad University(哈马碧大学)
专题命中
幻觉与鲁棒性
:VLM(summary_cn,abstract);vision language model(title,abstract);分类 cs.CV、cs.AI
机构
*
Hong Kong Polytechnic University(香港理工大学)
;
Nanyang Technological University(南洋理工大学)
;
Tsinghua University(清华大学)
;
National University of Singapore(新加坡国立大学)
专题命中
幻觉与鲁棒性
:MLLM(summary_cn,abstract_cn);multimodal large language model(title);分类 cs.CV、cs.AI、cs.LG
MirrorCheck: Efficient Adversarial Defense for Vision-Language Models
MirrorCheck: 视觉-语言模型的高效对抗防御
Samar Fares, Klea Ziu, Toluwani Aremu, Nikita Durasov, Martin Takáč, Pascal Fua, Ivan Laptev, Karthik Nandakumar
机构
*
Mohamed Bin Zayed University of Artificial Intelligence(莫扎伊德大学人工智能大学)
;
NVIDIA
;
École Polytechnique Fédérale de Lausanne(洛桑联邦理工学院)
;
Michigan State University(密歇根州立大学)
VOPE: Revisiting Hallucination of Vision-Language Models in Voluntary Imagination Task
VOPE:重新审视视觉语言模型在自愿想象任务中的幻觉现象
Xingming Long, Jie Zhang, Shiguang Shan, Xilin Chen
机构
*
Key Laboratory of AI Safety of CAS, Institute of Computing Technology, Chinese Academy of Sciences (CAS)(中国科学院人工智能安全重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Zhongguancun Academy(中关村学院)
6 Fingers, 1 Kidney: Natural Adversarial Medical Images Reveal Critical Weaknesses of Vision-Language Models
6根手指,1个肾脏:自然对抗性医学图像揭示视觉语言模型的关键弱点
Leon Mayer, Piotr Kalinowski, Caroline Ebersbach, Marcel Knopp, Tim Rädsch, Evangelia Christodoulou, Annika Reinke, Fiona R. Kolbinger, Lena Maier-Hein
机构
*
German Cancer Research Center (DKFZ) Heidelberg, Division of Intelligent Medical Systems(德国癌症研究中心(DKFZ)海德堡,智能医学系统部门)
;
Medical Faculty, Heidelberg University(海德堡大学医学院)
;
Faculty of Mathematics and Computer Science, Heidelberg University(海德堡大学数学与计算机科学学院)
;
HIDSS4Health - Helmholtz Information and Data Science School for Health, Karlsruhe/Heidelberg(HIDSS4Health - 哈勃-马克斯信息与数据科学健康学院,卡尔斯鲁厄/海德堡)
;
Helmholtz Imaging, German Cancer Research Center (DKFZ)(哈勃-马克斯成像,德国癌症研究中心(DKFZ))
;
Engineering Faculty, Heidelberg University(海德堡大学工程学院)
;
School of Computation, Information and Technology, TUM(技术大学(TUM)计算、信息与技术学院)
;
Weldon School of Biomedical Engineering, Purdue University(普渡大学韦尔登生物医学工程学院)
;
Department of Visceral, Thoracic and Vascular Surgery, University Hospital and Faculty of Medicine Carl Gustav Carus, TUD Dresden University of Technology(visceral、胸腔和血管外科部门,技术大学(TUD)德累斯顿大学医院和医学院)
;
National Center for Tumor Diseases (NCT), NCT Heidelberg, a partnership between DKFZ and University Hospital Heidelberg(肿瘤疾病国家中心(NCT),海德堡NCT,DKFZ与海德堡大学医院之间的合作)
;
Heidelberg University Hospital, Surgical Clinic, Surgical AI Research Group(海德堡大学医院,外科诊所,外科人工智能研究组)
;
Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI), Abu Dhabi, UAE(Mohamed Bin Zayed人工智能大学(MBZUAI),阿布扎赫,阿拉伯联合酋长国)
机构
*
Peking University(北京大学)
;
University of Pennsylvania(宾夕法尼亚大学)
;
Nanyang Technological University(南洋理工大学)
;
Tsinghua University(清华大学)
;
Virginia Tech(弗吉尼亚理工大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
De Artificial Intelligence Lab(人工智能实验室)