MIKO: Multimodal Intention Knowledge Distillation from Large Language Models for Social-Media Commonsense Discovery
专题命中 视觉定位与Grounding :multimodal large language model(abstract);MLLM(abstract)
Comments 11 pages, 5 figures
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :multimodal large language model(abstract);MLLM(abstract)
Comments 11 pages, 5 figures
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted to ICLR 2024, Project Page: http://ground-a-video.github.io
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted to AAAI 2024
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI、cs.LG
Comments version_02
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments EMNLP 2023
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted at the International Symposium on Experimental Robotics (ISER) 2023. Videos at http://kis-gmm.cs.uni-freiburg.de/
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Journal ref Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2023
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted at NeurIPS 2023 Track on Datasets and Benchmarks; Project Webpage: https://antoyang.github.io/vidchapters.html ; 31 pages; 8 figures
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments 14 pages, 5 figures
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted by MICCAI 2023
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments TLDR: Physical scenes are equivalence classes of sufficient statistics, and can be inferred uniquely by any agent measuring the same finite data; We formalize and implement an approach to representation learning that overturns "naive realism" in favor of an analytical approach of Russell and Koenderink. NeRFs cannot capture the physical scenes, but combined with Diffusion Models they can
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI、cs.LG
Comments CVPR 2023 (Highlighted Paper). Website: https://imagebind.metademolab.com/ Code/Models: https://github.com/facebookresearch/ImageBind
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Technical Report
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted: 1st Workshop on Safe Learning for Autonomous Driving, at the International Conference on Machine Learning (ICML 2022); Best Paper Award
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments NeurIPS 2022; updated with reviewers' comments addressed; Code is released at https://github.com/microsoft/GLIP
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted at NeurIPS 2022
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments CVPR 2022; updated visualizations; fixed hyper-parameters in Appendix C.1
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments paper accepted in CVPR 2022
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Project website at https://huangwl18.github.io/language-planner
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted as Student Abstract at AAAI-22
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted at Novel Ideas in Learning-to-Learn through Interaction (NILLI) workshop @ EMNLP 2021
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments NeurIPS 2021 (19 pages)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Project page: https://shivanshpatel35.github.io/comon/ ; the first three authors contributed equally
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG
Comments Tech Report. Authors Xin Zhou, Le Kang, and Zhiyu Cheng made equal contributions
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI、cs.LG