Challenges in Grounding Language in the Real World
专题命中 视觉定位与Grounding :grounding(title);分类 cs.AI
Comments 14 pages, 2 figures
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :grounding(title);分类 cs.AI
Comments 14 pages, 2 figures
机构 * National Key Laboratory of Human-Machine Hybrid Augmented Intelligence(国家人类-机器混合增强智能重点实验室) ; National Engineering Research Center for Visual Information and Applications(国家视觉信息与应用工程研究中心) ; Institute of Artificial Intelligence and Robotics(人工智能与机器人研究所) ; Xi’an Jiaotong University(西安交通大学)
专题命中 视觉定位与Grounding :grounding(title);分类 cs.CV
机构 * Institute Science of Tokyo(东京科学研究所)
专题命中 视觉定位与Grounding :vision-language model(title);分类 cs.CV
机构 * Mila-Québec, Université de Montréal(蒙特利尔大学魁北克分校) ; Carnegie Mellon University(卡内基梅隆大学) ; Deutsches Forschungszentrum für künstliche Intelligenz (DFKI)(德国人工智能研究中心) ; University of Pennsylvania(宾夕法尼亚大学) ; Next Generation Analytics and Modulo Bio(下一代分析与Modulo Bio) ; ServiceNow Research(ServiceNow研究) ; Mila-Québec, Université Laval(魁北克蒙特利尔大学拉瓦尔分校)
专题命中 视觉定位与Grounding :grounding(title);分类 cs.LG
Comments 39 pages, 8 figures; CLeaR 2025
机构 * Sun Yat-sen University, Guangzhou, China(中山大学)
专题命中 视觉定位与Grounding :grounding(title);分类 cs.AI
专题命中 视觉定位与Grounding :vision-language model(title);分类 cs.CV
Comments Accepted at ICIP2025 Dataset and Benchmark Track
机构 * Nanjing University(南京大学) ; Shanghai AI Laboratory(上海人工智能实验室) ; The Hong Kong University of Science and Technology(香港科学与技术大学) ; The Chinese University of Hong Kong(香港中文大学)
专题命中 视觉定位与Grounding :grounding(title);分类 cs.AI
Comments Accepted by ICML 2025, 21 pages
专题命中 视觉定位与Grounding :grounding(title);分类 cs.AI
Comments Presented at the 15th International Conference on Computational Creativity (ICCC'24)
Journal ref Proceedings of the Fifteenth International Conference on Computational Creativity (2024) 101-106
专题命中 视觉定位与Grounding :grounding(title);分类 cs.AI
Comments Oral at NAACL 2025 Main conference. Albuquerque, USA. Apr 29 - May 4, 2025. 19 pages, 9 figures, 7 tables
专题命中 视觉定位与Grounding :vision-language model(title);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title);分类 cs.CV
Comments 8 figures; 6 tables
专题命中 视觉定位与Grounding :grounding(title);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title);分类 cs.AI
专题命中 视觉定位与Grounding :grounding(title);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title);分类 cs.AI
Comments Add repair model ablation, update related work
专题命中 视觉定位与Grounding :grounding(title);分类 cs.CV
Comments Updated in Jan. 2025, In Proceedings of the European Conference on Computer Vision 2022 [ECCV 2022], 27 pages
专题命中 视觉定位与Grounding :grounding(title);分类 cs.AI
Comments Accepted at COLING 2025
专题命中 视觉定位与Grounding :multimodal large language model(title);分类 cs.AI
专题命中 视觉定位与Grounding :vision-language model(title);分类 cs.CV
Comments ICVGIP 2024, Young Faculty Symposium
专题命中 视觉定位与Grounding :vision language model(title);分类 cs.CV
Comments Accepted to the Thirty-Seventh Annual Conference on Innovative Applications of Artificial Intelligence (IAAI-25)
专题命中 视觉定位与Grounding :vision-language model(title);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title);分类 cs.CV
Comments Accepted by NeurIPS 2024
专题命中 视觉定位与Grounding :grounding(title);分类 cs.CV
Comments Preprint; 9 pages; 2024 EMNLP Findings
专题命中 视觉定位与Grounding :grounding(title);分类 cs.LG
Comments Accepted to EMNLP 2024 Findings
专题命中 视觉定位与Grounding :vision-language model(title);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title);分类 cs.CV
Comments This study was primarily conducted during the latter half of 2023
专题命中 视觉定位与Grounding :multimodal large language model(title);分类 cs.CV
专题命中 视觉定位与Grounding :multimodal large language model(title);分类 cs.AI
Comments To appear in eCrime 2024
专题命中 视觉定位与Grounding :grounding(title);分类 cs.CV
Comments ECCV 2024