arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 3460 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 跨模态检索 3460 篇

2008.01191 2020-08-05 cs.IR cs.CV cs.LG 57%

Deep Learning Techniques for Future Intelligent Cross-Media Retrieval

Sadaqat ur Rehman, Muhammad Waqas, Shanshan Tu, Anis Koubaa, Obaid ur Rehman, Jawad Ahmad, Muhammad Hanif, Zhu Han

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments arXiv admin note: text overlap with arXiv:1804.09539 by other authors

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.00949 2020-07-28 cs.LG cs.CL cs.DB stat.ML 57%

Attributed Sequence Embedding

Zhongfang Zhuang, Xiangnan Kong, Elke Rundensteiner, Jihane Zouaoui, Aditya Arora

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CL

Comments Accepted by IEEE Big Data 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.09464 2020-07-21 cs.CV cs.IR 57%

A Bag of Visual Words Model for Medical Image Retrieval

Sowmya Kamath S, Karthik K

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments In the proceedings of the 7th International Engineering Symposium (IES 2018), Kumamoto University, Kumamoto, Japan, Mar 7-9, 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.05028 2020-07-13 cs.LG cs.CV stat.ML 57%

Multi-view Orthonormalized Partial Least Squares: Regularizations and Deep Extensions

Li Wang, Ren-Cang Li, Wen-Wei

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.03375 2020-07-08 cs.CV 57%

Location Sensitive Image Retrieval and Tagging

Raul Gomez, Jaume Gibert, Lluis Gomez, Dimosthenis Karatzas

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Journal ref ECCV 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.09904 2020-06-18 cs.IR cs.CV cs.LG 57%

Learning Colour Representations of Search Queries

Paridhi Maheshwari, Manoj Ghuhan, Vishwa Vinay

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted as a full paper at SIGIR 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.14169 2020-05-29 cs.CV 57%

Self-supervised Modal and View Invariant Feature Learning

Longlong Jing, Yucheng Chen, Ling Zhang, Mingyi He, Yingli Tian

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.08399 2020-05-19 cs.CV 57%

T-VSE: Transformer-Based Visual Semantic Embedding

Muhammet Bastan, Arnau Ramisa, Mehmet Tek

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments To appear: CVPR 2020 Workshop on Computer Vision for Fashion, Art and Design (CVFAD 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.10310 2020-05-13 cs.CV 57%

Sketch Less for More: On-the-Fly Fine-Grained Sketch Based Image Retrieval

Ayan Kumar Bhunia, Yongxin Yang, Timothy M. Hospedales, Tao Xiang, Yi-Zhe Song

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments IEEE Conf. on Computer Vision and Pattern Recognition (CVPR), 2020 [Oral Presentation] Code: https://github.com/AyanKumarBhunia/on-the-fly-FGSBIR

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.02133 2020-05-06 cs.CV eess.IV stat.ML 57%

Quality Guided Sketch-to-Photo Image Synthesis

Uche Osahor, Hadi Kazemi, Ali Dabouei, Nasser Nasrabadi

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments 10 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.00725 2020-05-05 cs.CV 57%

Cross-View Image Retrieval -- Ground to Aerial Image Retrieval through Deep Learning

Numan Khurshid, Talha Hanif, Mohbat Tharani, Murtaza Taj

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments International Conference on Neural Information Processing (ICONIP-2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.11892 2020-04-07 cs.CV cs.LG stat.ML 57%

CLAREL: Classification via retrieval loss for zero-shot learning

Boris N. Oreshkin, Negar Rostamzadeh, Pedro O. Pinheiro, Christopher Pal

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.12299 2020-03-31 cs.CV 57%

CurlingNet: Compositional Learning between Images and Text for Fashion IQ Data

Youngjae Yu, Seunghwan Lee, Yuncheol Choi, Gunhee Kim

专题命中 跨模态检索 :image-text(abstract);分类 cs.CV

Comments 4 pages, 4 figures, ICCV 2019 Linguistics Meets image and video retrieval workshop, Fashion IQ challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.00636 2020-03-03 cs.CV cs.LG eess.IV 57%

Matching Neuromorphic Events and Color Images via Adversarial Learning

Fang Xu, Shijie Lin, Wen Yang, Lei Yu, Dengxin Dai, Gui-song Xia

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.05493 2020-01-22 cs.CL cs.IR 57%

A Unified System for Aggression Identification in English Code-Mixed and Uni-Lingual Texts

Anant Khandelwal, Niraj Kumar

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CL

Comments 10 pages, 5 Figures, 6 Tables, accepted at CoDS-COMAD 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.04732 2020-01-15 cs.CV 57%

Fine-grained Image Classification and Retrieval by Combining Visual and Locally Pooled Textual Features

Andres Mafla, Sounak Dey, Ali Furkan Biten, Lluis Gomez, Dimosthenis Karatzas

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments Winter Conference on Applications of Computer Vision (WACV 2020) Accepted paper

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.08904 2019-12-20 cs.IR cs.CL cs.HC 57%

Macaw: An Extensible Conversational Information Seeking Platform

Hamed Zamani, Nick Craswell

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.01770 2019-11-06 cs.IR cs.CL 57%

Self-Attention and Ingredient-Attention Based Model for Recipe Retrieval from Image Queries

Matthias Fontanellaz, Stergios Christodoulidis, Stavroula Mougiakakou

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CL

Comments MADiMa 2019,5th International Workshop on Multimedia Assisted Dietary Management

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.11150 2019-10-22 eess.IV cs.CV 57%

Hetero-Modal Variational Encoder-Decoder for Joint Modality Completion and Segmentation

Reuben Dorent, Samuel Joutard, Marc Modat, Sébastien Ourselin, Tom Vercauteren

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments Accepted at MICCAI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.02327 2019-10-21 cs.CV 57%

Semantic Adversarial Network for Zero-Shot Sketch-Based Image Retrieval

Xinxun Xu, Hao Wang, Leida Li, Cheng Deng

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments There is a big problem with the paper and I hope it can be retracted

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.11393 2019-09-25 cs.CL 57%

Learning semantic sentence representations from visually grounded language without lexical knowledge

Danny Merkx, Stefan Frank

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CL

Journal ref Natural Language Engineering, Volume 25 - Issue 4 - July 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.01976 2019-09-06 cs.CV 57%

Do Cross Modal Systems Leverage Semantic Relationships?

Shah Nawaz, Muhammad Kamran Janjua, Ignazio Gallo, Arif Mahmood, Alessandro Calefati, Faisal Shafait

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted to cross modal learning in real world in conjunction with ICCV 2019. arXiv admin note: text overlap with arXiv:1807.07364

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.03704 2019-09-06 cs.CV 57%

SMIT: Stochastic Multi-Label Image-to-Image Translation

Andrés Romero, Pablo Arbeláez, Luc Van Gool, Radu Timofte

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments ICCV Workshops, 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.02950 2019-08-09 cs.CV eess.IV 57%

Semi Supervised Phrase Localization in a Bidirectional Caption-Image Retrieval Framework

Deepan Das, Noor Mohammed Ghouse, Shashank Verma, Yin Li

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.10428 2019-07-25 cs.LG cs.HC cs.SD eess.AS 57%

EmoBed: Strengthening Monomodal Emotion Recognition via Training with Crossmodal Emotion Embeddings

Jing Han, Zixing Zhang, Zhao Ren, Björn Schuller

专题命中 跨模态检索 :multimodal(abstract);分类 eess.AS

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.09775 2019-07-24 cs.RO cs.CV cs.SD 57%

Multisensory Learning Framework for Robot Drumming

A. Barsky, C. Zito, H. Mori, T. Ogata, J. L. Wyatt

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Extended abstract

Journal ref Workshop on Crossmodal Learning for Intelligent Robotics 2nd Edition. IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.06011 2019-07-03 cs.CV 57%

Fusion vectors: Embedding Graph Fusions for Efficient Unsupervised Rank Aggregation

Icaro Cavalcante Dourado, Ricardo da Silva Torres

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.12254 2019-04-30 cs.CV 57%

Translate-to-Recognize Networks for RGB-D Scene Recognition

Dapeng Du, Limin Wang, Huiling Wang, Kai Zhao, Gangshan Wu

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted by CVPR 2019. Project: https://ownstyledu.github.io/Translate-to-Recognize-Networks/

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.04985 2019-04-11 cs.CV 57%

Context-Aware Embeddings for Automatic Art Analysis

Noa Garcia, Benjamin Renoust, Yuta Nakashima

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.04272 2019-04-10 cs.LG cs.CV cs.IR stat.ML 57%

SoDeep: a Sorting Deep net to learn ranking loss surrogates

Martin Engilberge, Louis Chevallier, Patrick Pérez, Matthieu Cord

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted to CVPR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏