机构
*
University of Michigan(密歇根大学)
;
The Ohio State University(俄亥俄州立大学)
;
Independent(独立)
;
Indiana University(印第安纳大学)
;
The University of Hong Kong(香港大学)
;
Peking University(北京大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
LMSYS Org(LMSYS组织)
Multi-PA: A Multi-perspective Benchmark on Privacy Assessment for Large Vision-Language Models
多视角:面向大视觉-语言模型隐私评估的基准测试
Jie Zhang, Xiangkui Cao, Zhouyu Han, Shiguang Shan, Xilin Chen
机构
*
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(人工智能安全国家重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Anhui Province Key Laboratory of Digital Security(安徽省数字安全重点实验室)
;
National University of Singapore(新加坡国立大学)
Causality-guided Prompt Learning for Vision-language Models via Visual Granulation
基于视觉粒化的因果引导提示学习用于视觉语言模型
Mengyu Gao, Qiulei Dong
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学)
机构
*
Department of Computing, The Hong Kong Polytechnic University(香港理工大学计算机系)
;
Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative, Institute of Digital Twin, EIT(宁波空间智能与数字衍生关键实验室,数字孪生研究院,EIT)
;
Shanghai Jiao Tong University(上海交通大学)
;
Ocean University of China(中国海洋大学)
专题命中
其他VLM
:multimodal large language model(title,abstract);分类 cs.CV
MetaTPT: Meta Test-time Prompt Tuning for Vision-Language Models
MetaTPT: 用于视觉-语言模型的元测试时间提示微调
Yuqing Lei, Yingjun Du, Yawen Huang, Xiantong Zhen, Ling Shao
机构
*
UCAS-Terminus AI Lab, University of Chinese Academic of Sciences(中国科学院大学Terminus AI实验室,中国科学院大学)
;
AIM Lab, University of Amsterdam(阿姆斯特丹大学AIM实验室)
;
Jarvis Research Center, Tencent Youtu Lab(腾讯优图实验室 Jarvis 研究中心)
;
Central Research Institue, United Imaging Healthcare Co., Ltd(联合影像医疗科技有限公司中央研究所)
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学北京校区人工智能学院)
;
Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大型模型与智能治理研究重点实验室)
;
Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程研究中心)
;
Dept. of Comp. Sci. & Tech., Tsinghua University(清华大学计算机科学与技术系)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
;
SKL-ESPC & SEPKL-AERM, College of Environmental Sciences and Engineering, Peking University(环境科学与工程学院,北京大学)
;
MAIS, Institute of Automation, Chinese Academy of Sciences(自动化研究所,中国科学院)
;
DAMO Academy, Alibaba Group, Hangzhou, China(阿里巴巴集团 DAMO Academy,中国杭州)
;
Hupan Lab, Hangzhou, China(杭州实验室)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Nanyang Technological University(南洋理工大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Sichuan University(四川大学)
;
National University of Singapore(新加坡国立大学)
;
Shenzhen University(深圳大学)
The Potential and Limitations of Vision-Language Models for Human Motion Understanding: A Case Study in Data-Driven Stroke Rehabilitation
视觉-语言模型在人类运动理解中的潜力与局限性:数据驱动中风康复的案例研究
Victor Li, Naveenraj Kamalakannan, Avinash Parnandi, Heidi Schambra, Carlos Fernandez-Granda
机构
*
Center for Data Science(数据科学中心)
;
Tandon School of Engineering(工程学院)
;
VitalConnect(VitalConnect公司)
;
Department of Neurology(神经病学部)
;
Department of Rehabilitation Medicine(康复医学部)
;
Courant Institute of Mathematical Sciences(数学科学研究所)
机构
*
State Key Laboratory of General Artificial Intelligence, Beijing Institute for General Artificial Intelligence(通用人工智能国家重点实验室、北京通用人工智能研究院)
;
School of Psychological and Cognitive Sciences and Beijing Key Laboratory of Behavior and Mental Health, Key Laboratory of Machine Perception (Ministry of Education), Peking University(心理与认知科学学院及北京行为与心理健康重点实验室、机器感知重点实验室(教育部))
;
School of Intelligence Science and Technology, Peking University(智能科学与技术学院)
专题命中
其他VLM
:multimodal large language model(title,abstract);分类 cs.AI