Evaluating Linguistic Capabilities of Multimodal LLMs in the Lens of Few-Shot Learning
专题命中 视觉定位与Grounding :grounding(abstract);multimodal large language model(abstract);分类 cs.CV
Comments Preprint. 33 pages, 17 Figures, 3 Tables
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :grounding(abstract);multimodal large language model(abstract);分类 cs.CV
Comments Preprint. 33 pages, 17 Figures, 3 Tables
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
Comments Accepted at ECCV 2024
专题命中 视觉定位与Grounding :vision language model(abstract);grounding(abstract);分类 cs.CV
Journal ref ECCV 2024
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
Comments 4 pages
专题命中 视觉定位与Grounding :grounding(abstract);multimodal large language model(abstract);分类 cs.CV
Comments Accepted at ECCV 2024
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
Comments Technical report. Extension of CVPR paper "Open-vocabulary object 6D pose estimation". Project page: https://jcorsetti.github.io/oryon
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
Comments Update details. Acceptd by ECCV 2024
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :multimodal large language model(abstract);MLLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision language model(abstract);grounding(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision language model(abstract);grounding(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(abstract);grounding(abstract);分类 cs.CV
Comments CVPR 2024
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
Comments Accepted at NAACL 2024
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
Comments CVPR 2024. Project website at https://lucazanella.github.io/lavad/
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision language model(abstract);VLM(abstract);分类 cs.CV
Comments 7 pages, 7 figures. Accepted to IEEE International Conference on Robotics and Automation (ICRA) 2024
专题命中 视觉定位与Grounding :vision language model(abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(abstract);grounding(abstract);分类 cs.CV
Comments accepted to NeurIPS 2023
专题命中 视觉定位与Grounding :vision-language model(abstract);grounding(abstract);分类 cs.CV
Comments Project page: https://jerryxu.net/PixelLLM
专题命中 视觉定位与Grounding :vision-language model(abstract);grounding(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :multimodal large language model(abstract);MLLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :visual language model(abstract);grounding(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
Comments Accepted by NeurIPS 2023
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :visual reasoning(abstract);grounding(abstract);分类 cs.CV
Comments Camera-ready version, code is in https://github.com/sail-sg/ptp
专题命中 视觉定位与Grounding :vision-language model(abstract);grounding(abstract);分类 cs.CV
Comments Accepted to CVPR2023
专题命中 视觉定位与Grounding :vision-language model(abstract);grounding(abstract);分类 cs.CV
Comments Project website can be found at https://lerf.io