Efficient Open Set Single Image Test Time Adaptation of Vision Language Models
专题命中 其他VLM :vision language model(title);vision-language model(abstract);分类 cs.CV
Comments Accepted at TMLR
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 其他VLM :vision language model(title);vision-language model(abstract);分类 cs.CV
Comments Accepted at TMLR
机构 * MIT(麻省理工学院) ; Georgia Tech(佐治亚理工学院) ; Meta ; UC Davis(加州大学戴维斯分校)
专题命中 其他VLM :vision language model(title,abstract);分类 cs.CV
Comments Annual Meeting of the Association for Computational Linguistics (ACL), 2025
机构 * Pengfei Wang(王鹏飞) ; Guohai Xu(徐国海) ; Weinong Wang(王文龙) ; Junjie Yang(杨俊杰) ; Jie Lou(娄杰) ; Yunhua Xue(许云华)
专题命中 其他VLM :multimodal large language model(title,abstract);分类 cs.CV
机构 * Robotics Program KAIST(韩国釜山科学技术院机器人计划) ; Electrical Engineering KAIST(韩国釜山科学技术院电子工程)
专题命中 其他VLM :vision language model(title);vision-language model(abstract);分类 cs.LG
Comments 5 pages, ICASSP 2025. The first two authors are equally contributed
Journal ref ICASSP 2025 - 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
机构 * Institution1(机构1) ; Institution2(机构2)
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
Comments CVPR 2025; project page: https://negbench.github.io
机构 * Information & Computer Science Department, King Fahd University of Petroleum & Minerals(国王法赫德石油与矿物大学信息与计算机科学系) ; Center for Machine Vision and Signal Analysis, University of Oulu(奥卢大学机器视觉与信号分析中心)
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :multimodal large language model(title);MLLM(abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
Comments CVPR 2025
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
Comments Accepted to CVPR 2025
专题命中 其他VLM :MLLM(title);multimodal large language model(abstract);分类 cs.AI
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.AI
Comments Under review at npj Digital Medicine
专题命中 其他VLM :multimodal large language model(title,abstract);分类 cs.AI
Comments NAACL Main 2025
专题命中 其他VLM :multimodal large language model(title,abstract);分类 cs.AI
Comments revised version 2. September 2024
专题命中 其他VLM :MLLM(title);multimodal large language model(abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.AI
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.LG
Comments 18 pages, 5 figures, 13 tables, GitHub repository: https://github.com/Seefreem/meme_text_retrieval_p1
专题命中 其他VLM :vision-language model(title);visual language model(abstract);分类 cs.CV
Comments AAAI(Oral)
专题命中 其他VLM :multimodal large language model(title,abstract);分类 cs.CV
Comments Accepted by the AAAI 2025
专题命中 其他VLM :multimodal large language model(title);MLLM(abstract);分类 cs.CV
Comments Accepted to IEEE Transactions on Multimedia (TMM)
专题命中 其他VLM :visual language model(title);vision-language model(abstract);分类 cs.AI
Comments Updated version of the paper presented in 2025 AIAA SciTech. https://arc.aiaa.org/doi/10.2514/6.2025-1543
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(title,abstract);分类 cs.CV
Comments Accepted by AAAI2025