Can Multimodal Large Language Models be Guided to Improve Industrial Anomaly Detection?
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
Comments 16 pages, 11 figures
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
Comments 16 pages, 11 figures
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
Comments 10 pages
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
Comments Accepted by AAAI 2025
专题命中 视觉定位与Grounding :VLM(title);vision language model(abstract);visual reasoning(abstract);分类 cs.CV
Comments Accepted for publication at aaai25, project page: https://jasonjin34.github.io/logicad.github.io/
专题命中 视觉定位与Grounding :VLM(title,abstract);vision-language model(abstract);分类 cs.CV
Comments Accepted by NeurIPS 2024
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
Comments 13 pages, 10 figures
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
Comments Accepted as Proceedings Paper at ML4H 2024
专题命中 视觉定位与Grounding :VLM(title,abstract);vision-language model(abstract);分类 cs.CV
Comments WACV 2025 Accepted
专题命中 视觉定位与Grounding :vision language model(title,abstract);LLaVA(abstract);分类 cs.CV
Comments 52 pages, 13 figures
专题命中 视觉定位与Grounding :grounding(title,abstract);visual language model(abstract);分类 cs.CV
Comments Accepted by NeurIPS 2024. The project page: https://github.com/linhuixiao/OneRef
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);grounding(abstract);分类 cs.CV
Comments 8th Conference on Robot Learning (CoRL 2024), Munich, Germany
专题命中 视觉定位与Grounding :vision language model(title,abstract);VLM(abstract);分类 cs.CV
Comments IEEE Robotics and Automation Letters
专题命中 视觉定位与Grounding :grounding(title,abstract);vision language model(abstract);分类 cs.AI
Comments Published in IROS 2024
专题命中 视觉定位与Grounding :grounding(title,abstract);vision-language model(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :visual language model(title,abstract);grounding(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title,abstract);visual language model(abstract);分类 cs.CV
Comments Accepted by ECCV 2024
专题命中 视觉定位与Grounding :vision language model(title,abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title,abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
Comments 29 pages, Accepted for publication in ECCV 2024
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
Comments CVPR 2024 Accepted
专题命中 视觉定位与Grounding :grounding(title,abstract);visual question answering(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
Comments Accepted to NAACL 2024 Main Conference
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :grounding(title,abstract);visual question answering(abstract);分类 cs.CV