INQUIRE: A Natural World Text-to-Image Retrieval Benchmark
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments Published in NeurIPS 2024, Datasets and Benchmarks Track
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments Published in NeurIPS 2024, Datasets and Benchmarks Track
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted to WACV 2025
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV、cs.AI、cs.MM
Comments Work in progress
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments CODE: https://github.com/UX-Decoder/FIND
专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 跨模态检索 :image-text(abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 跨模态检索 :multimodal(abstract);cross-modal(abstract)
Comments ICLR 2024
专题命中 跨模态检索 :multi-modal(abstract);cross-modal(abstract)
Comments ICDE'24 Accepted
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments 17 pages (Our code is available at https://github.com/adymaharana/d2pruning)
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV、cs.CL、cs.MM
Comments arXiv admin note: text overlap with arXiv:2210.09338 by other authors
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV、cs.AI、cs.MM
Comments CVPR 2023 (Highlighted Paper). Website: https://imagebind.metademolab.com/ Code/Models: https://github.com/facebookresearch/ImageBind
专题命中 跨模态检索 :multimodal(abstract);cross-modal(abstract)
Comments 5 pages, accepted to The Industry Track of the Web Conference 2023
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments Published at ICLR 2023
专题命中 跨模态检索 :image-text(abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.AI、cs.MM
专题命中 跨模态检索 :image-text(abstract);分类 cs.CV、cs.CL、cs.AI
Comments 5 pages, 2 figures, accepted by 2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP 2022)
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted as Student Abstract at AAAI-22
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted for publication at the 1st Multilingual Representation Learning workshop (MRL 2021) co-located with EMNLP 2021. 15 pages, 8 figures, 6 tables
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.AI、cs.MM
Comments 17 pages, 3 figures
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 跨模态检索 :multi-modal(abstract);cross-modal(abstract)
Comments 13 Pages, 6 Figures and 10 Tables
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.CL、cs.MM
Comments Accepted to ACM Transactions on Multimedia Computing Communications and Applications (ACM TOMM)
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments 3 pages including references, Accepted at the ICCV 2019 Workshop - 'Linguistics Meets Image and Video Retrieval' (received Best Paper Award)
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted at WACV 2019. Also at NeurIPS 2017 workshop on Visually-Grounded Interaction and Language (ViGIL)
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.AI、cs.MM
Comments *Ayush Jaiswal and Ekraam Sabir contributed equally to the work in this paper
Journal ref In Proceedings of the 2017 ACM on Multimedia Conference, pp. 1465-1471. ACM, 2017
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted to BMVC'16