arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 3460 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 跨模态检索 3460 篇

1904.03223 2019-04-09 cs.CL cs.IR 57%

NELEC at SemEval-2019 Task 3: Think Twice Before Going Deep

Parag Agrawal, Anshuman Suri

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CL

Comments International Workshop on Semantic Evaluation (SemEval), NAACL-HLT 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.05743 2019-03-25 cs.IR cs.CV 57%

Unsupervised Graph-based Rank Aggregation for Improved Retrieval

Icaro Cavalcante Dourado, Daniel Carlos Guimarães Pedronette, Ricardo da Silva Torres

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.08863 2019-03-22 cs.CV cs.LG 57%

Learning Disentangled Representations of Satellite Image Time Series

Eduardo Sanchez, Mathieu Serrurier, Mathias Ortner

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.08896 2019-03-19 cs.CV 57%

Towards Practical Visual Search Engine within Elasticsearch

Cun Mu, Jun Zhao, Guang Yang, Jing Zhang, Zheng Yan

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments Accepted by SIGIR eCom'18

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.04715 2019-03-13 cs.CL 57%

Context-Aware Learning for Neural Machine Translation

Sébastien Jean, Kyunghyun Cho

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.03878 2019-02-12 cs.MM cs.IR 57%

Towards an All-Purpose Content-Based Multimedia Information Retrieval System

Ralph Gasser, Luca Rossetto, Heiko Schuldt

专题命中 跨模态检索 :multimodal(abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.09650 2018-12-27 cs.CL cs.LG stat.ML 57%

Improving Context-Aware Semantic Relationships in Sparse Mobile Datasets

Peter Hansel, Nik Marda, William Yin

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CL

Comments 6 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.09730 2018-11-06 cs.CV 57%

Image-to-image translation for cross-domain disentanglement

Abel Gonzalez-Garcia, Joost van de Weijer, Yoshua Bengio

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments Accepted to NIPS 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.10226 2018-10-25 cs.IR cs.LG cs.MM 57%

Textually Guided Ranking Network for Attentional Image Retweet Modeling

Zhou Zhao, Hanbing Zhan, Lingtao Meng, Jun Xiao, Jun Yu, Min Yang, Fei Wu, Deng Cai

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.MM

Comments 12 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.05267 2018-09-17 cs.RO cs.CV 57%

Detection-by-Localization: Maintenance-Free Change Object Detector

Tanaka Kanji

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments 7 pages, 3 figures, Technical report

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.01362 2018-07-17 cs.CV 57%

Predicting Visual Features from Text for Image and Video Caption Retrieval

Jianfeng Dong, Xirong Li, Cees G. M. Snoek

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments Accepted by Transaction on Multimedia. Code is available at https://github.com/danieljf24/w2vv

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.02110 2018-07-09 cs.CV 57%

TextTopicNet - Self-Supervised Learning of Visual Features Through Embedding Images on Semantic Text Spaces

Yash Patel, Lluis Gomez, Raul Gomez, Marçal Rusiñol, Dimosthenis Karatzas, C. V. Jawahar

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments arXiv admin note: text overlap with arXiv:1705.08631

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.00974 2018-06-06 cs.CV 57%

ALMN: Deep Embedding Learning with Geometrical Virtual Point Generating

Binghui Chen, Weihong Deng

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1712.00981 2018-04-13 cs.CV 57%

Feature Generating Networks for Zero-Shot Learning

Yongqin Xian, Tobias Lorenz, Bernt Schiele, Zeynep Akata

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments 2018 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.02284 2018-03-07 cs.CV 57%

Zero-Shot Sketch-Image Hashing

Yuming Shen, Li Liu, Fumin Shen, Ling Shao

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted as spotlight at CVPR 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.04154 2017-11-15 cs.CL 57%

Interpretable probabilistic embeddings: bridging the gap between topic models and neural networks

Anna Potapenko, Artem Popov, Konstantin Vorontsov

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CL

Comments Appeared in AINL-2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.07534 2017-09-25 cs.AI cs.LG stat.ML 57%

MRNet-Product2Vec: A Multi-task Recurrent Neural Network for Product Embeddings

Arijit Biswas, Mukul Bhutani, Subhajit Sanyal

专题命中 跨模态检索 :multimodal(abstract);分类 cs.AI

Comments Published in ECML-PKDD 2017 (Applied Data Science Track)

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.06761 2017-09-04 cs.CV 57%

Content-Based Video-Music Retrieval Using Soft Intra-Modal Structure Constraint

Sungeun Hong, Woobin Im, Hyun S. Yang

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments 13 pages, 9 figures, 4 tables, supplementary material link >> https://youtu.be/ZyINqDMo3Fg

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.07945 2017-08-24 cs.CV 57%

Spatio-temporal Person Retrieval via Natural Language Queries

Masataka Yamaguchi, Kuniaki Saito, Yoshitaka Ushiku, Tatsuya Harada

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments Accepted to ICCV2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.08472 2017-08-02 cs.CV 57%

Medical Image Retrieval using Deep Convolutional Neural Network

Adnan Qayyum, Syed Muhammad Anwar, Muhammad Awais, Muhammad Majid

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments Submitted to Neurocomputing

Journal ref Neurocomputing 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.05691 2017-07-19 cs.CV 57%

Learning Fashion Compatibility with Bidirectional LSTMs

Xintong Han, Zuxuan Wu, Yu-Gang Jiang, Larry S. Davis

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments ACM MM 17

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.00932 2017-06-06 cs.CV 57%

See, Hear, and Read: Deep Aligned Representations

Yusuf Aytar, Carl Vondrick, Antonio Torralba

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.08631 2017-05-25 cs.CV 57%

Self-supervised learning of visual features through embedding images into text topic spaces

Lluis Gomez, Yash Patel, Marçal Rusiñol, Dimosthenis Karatzas, C. V. Jawahar

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments Accepted CVPR 2017 paper

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.07692 2017-05-23 cs.CV 57%

Semantic Softmax Loss for Zero-Shot Learning

Zhong Ji, Yunxin Sun, Yulong Yu, Jichang Guo, Yanwei Pang

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.04964 2017-05-16 cs.CV 57%

Machine learning methods for multimedia information retrieval

Bálint Zoltán Daróczy

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments doctoral thesis, 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.01958 2017-04-28 cs.CV 57%

Learning Diverse Image Colorization

Aditya Deshpande, Jiajun Lu, Mao-Chuang Yeh, Min Jin Chong, David Forsyth

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments This revision to appear in CVPR17

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.05539 2017-04-20 cs.AI 57%

Beating Atari with Natural Language Guided Reinforcement Learning

Russell Kaplan, Christopher Sauer, Alexander Sosa

专题命中 跨模态检索 :multimodal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.04394 2017-04-17 cs.CV 57%

DESIRE: Distant Future Prediction in Dynamic Scenes with Interacting Agents

Namhoon Lee, Wongun Choi, Paul Vernaza, Christopher B. Choy, Philip H. S. Torr, Manmohan Chandraker

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments Accepted at CVPR 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.07571 2017-04-12 cs.CV cs.LG cs.NE 57%

Quad-networks: unsupervised learning to rank for interest point detection

Nikolay Savinov, Akihito Seki, Lubor Ladicky, Torsten Sattler, Marc Pollefeys

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted at CVPR 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.03567 2017-03-13 cs.CV 57%

A New Evaluation Protocol and Benchmarking Results for Extendable Cross-media Retrieval

Ruoyu Liu, Yao Zhao, Liang Zheng, Shikui Wei, Yi Yang

专题命中 跨模态检索 :image-text(abstract);分类 cs.CV

Comments 10 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏