arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 3460 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 跨模态检索 3460 篇

2109.10477 2021-09-23 cs.CV cs.IR cs.LG 57%

Generating Compositional Color Representations from Text

Paridhi Maheshwari, Nihal Jain, Praneetha Vaddamanu, Dhananjay Raut, Shraiysh Vaishay, Vishwa Vinay

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted as a full paper at CIKM 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.05534 2021-09-14 cs.CV 57%

DSSL: Deep Surroundings-person Separation Learning for Text-based Person Retrieval

Aichun Zhu, Zijie Wang, Yifeng Li, Xili Wan, Jing Jin, Tian Wang, Fangqiang Hu, Gang Hua

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments Accepted by ACM MM'21

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.13592 2021-09-06 cs.IR cs.AI 57%

Zero Shot on the Cold-Start Problem: Model-Agnostic Interest Learning for Recommender Systems

Philip J. Feng, Pingjun Pan, Tingting Zhou, Hongxiang Chen, Chuanjiang Luo

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.00120 2021-09-03 cs.CV 57%

Contrastive Multiview Coding with Electro-optics for SAR Semantic Segmentation

Keumgang Cha, Junghoon Seo, Yeji Choi

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments To be appeared in IEEE GRSL. DOI to be updated

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.09305 2021-08-24 cs.LG cs.AI cs.CR 57%

Data-driven Smart Ponzi Scheme Detection

Yuzhi Liang, Weijing Wu, Kai Lei, Feiyang Wang

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.07353 2021-08-18 cs.CV 57%

Scene Designer: a Unified Model for Scene Search and Synthesis from Sketch

Leo Sampaio Ferraz Ribeiro, Tu Bui, John Collomosse, Moacir Ponti

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted to the 1st Workshop on Sketching for Human Expressivity (SHE), at ICCV 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.03914 2021-08-10 cs.MM 57%

Two-pronged Strategy: Lightweight Augmented Graph Network Hashing for Scalable Image Retrieval

Hui Cui, Lei Zhu, Jingjing Li, Zhiyong Cheng, Zheng Zhang

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.15307 2021-08-02 eess.SP cs.AI stat.ML 57%

Deep Random Projection Outlyingness for Unsupervised Anomaly Detection

Martin Bauw, Santiago Velasco-Forero, Jesus Angulo, Claude Adnet, Olivier Airiau

专题命中 跨模态检索 :multimodal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.00774 2021-07-09 cs.IR cs.AI cs.LG 57%

Fast Multi-Step Critiquing for VAE-based Recommender Systems

Diego Antognini, Boi Faltings

专题命中 跨模态检索 :multimodal(abstract);分类 cs.AI

Comments Accepted at RecSys 2021. 19 pages, 7 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.03671 2021-07-06 cs.CV cs.LG 57%

A Self-Supervised Gait Encoding Approach with Locality-Awareness for 3D Skeleton Based Person Re-Identification

Haocong Rao, Siqi Wang, Xiping Hu, Mingkui Tan, Yi Guo, Jun Cheng, Xinwang Liu, Bin Hu

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Pattern Analysis and Machine Intelligence (T-PAMI). Journal version of https://www.ijcai.org/proceedings/2020/0125 (IJCAI 2020). Codes are available at https://github.com/Kali-Hac/Locality-Awareness-SGE. arXiv admin note: text overlap with arXiv:2008.09435

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.00708 2021-07-05 cs.CV 57%

Blind Image Super-Resolution via Contrastive Representation Learning

Jiahui Zhang, Shijian Lu, Fangneng Zhan, Yingchen Yu

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.11841 2021-06-23 cs.CV 57%

Domain-Smoothing Network for Zero-Shot Sketch-Based Image Retrieval

Zhipeng Wang, Hao Wang, Jiexi Yan, Aming Wu, Cheng Deng

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted to IJCAI 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.01150 2021-05-10 cs.CL stat.AP 57%

Modeling Social Readers: Novel Tools for Addressing Reception from Online Book Reviews

Pavan Holur, Shadi Shahsavari, Ehsan Ebrahimzadeh, Timothy R. Tangherlini, Vwani Roychowdhury

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.14994 2021-05-05 cs.IR cs.MM 57%

GeoWINE: Geolocation based Wiki, Image,News and Event Retrieval

Golsa Tahmasebzadeh, Endri Kacupaj, Eric Müller-Budack, Sherzod Hakimov, Jens Lehmann, Ralph Ewerth

专题命中 跨模态检索 :multimodal(abstract);分类 cs.MM

Comments Accepted for publication in: International ACM SIGIR Conference on Research and Development in Information Retrieval 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.09918 2021-04-22 cs.CV cs.LG 57%

CrossATNet - A Novel Cross-Attention Based Framework for Sketch-Based Image Retrieval

Ushasi Chaudhuri, Biplab Banerjee, Avik Bhattacharya, Mihai Datcu

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted in Journal of Image and Vision Computing

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.08179 2021-04-13 cs.AI cs.LG 57%

DeepEnroll: Patient-Trial Matching with Deep Embedding and Entailment Prediction

Xingyao Zhang, Cao Xiao, Lucas M. Glass, Jimeng Sun

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.AI

Comments accepted by The World Wide Web Conference 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.01552 2021-04-06 cs.CV 57%

Scene Text Retrieval via Joint Text Detection and Similarity Learning

Hao Wang, Xiang Bai, Mingkun Yang, Shenggao Zhu, Jing Wang, Wenyu Liu

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted to CVPR 2021. Code is available at: https://github.com/lanfeng4659/STR-TDSL

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.15706 2021-04-01 cs.CV 57%

StyleMeUp: Towards Style-Agnostic Sketch-Based Image Retrieval

Aneeshan Sain, Ayan Kumar Bhunia, Yongxin Yang, Tao Xiang, Yi-Zhe Song

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments IEEE Conf. on Computer Vision and Pattern Recognition (CVPR), 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.08674 2021-03-30 cs.LG cs.CV 57%

Wasserstein Contrastive Representation Distillation

Liqun Chen, Dong Wang, Zhe Gan, Jingjing Liu, Ricardo Henao, Lawrence Carin

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted by CVPR 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.13990 2021-03-26 cs.CV 57%

More Photos are All You Need: Semi-Supervised Learning for Fine-Grained Sketch Based Image Retrieval

Ayan Kumar Bhunia, Pinaki Nath Chowdhury, Aneeshan Sain, Yongxin Yang, Tao Xiang, Yi-Zhe Song

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments IEEE Conf. on Computer Vision and Pattern Recognition (CVPR), 2021 Code : https://github.com/AyanKumarBhunia/semisupervised-FGSBIR

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.13942 2021-03-26 cs.CL 57%

Visual Grounding Strategies for Text-Only Natural Language Processing

Damien Sileo

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CL

Comments Accepted at LANTERN2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.10801 2021-01-27 cs.CV 57%

Global-Local Propagation Network for RGB-D Semantic Segmentation

Sihan Chen, Xinxin Zhu, Wei Liu, Xingjian He, Jing Liu

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.13089 2020-12-25 cs.CV 57%

P4Contrast: Contrastive Learning with Pairs of Point-Pixel Pairs for RGB-D Scene Understanding

Yunze Liu, Li Yi, Shanghang Zhang, Qingnan Fan, Thomas Funkhouser, Hao Dong

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.01045 2020-11-30 eess.IV cs.CV 57%

Brain tumor segmentation with self-ensembled, deeply-supervised 3D U-net neural networks: a BraTS 2020 challenge solution

Theophraste Henry, Alexandre Carre, Marvin Lerousseau, Theo Estienne, Charlotte Robert, Nikos Paragios, Eric Deutsch

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV

Comments BraTS 2020 proceedings (LNCS) paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.02779 2020-11-11 cs.CL 57%

UPB at SemEval-2020 Task 8: Joint Textual and Visual Modeling in a Multi-Task Learning Architecture for Memotion Analysis

George-Alexandru Vlad, George-Eduard Zaharia, Dumitru-Clementin Cercel, Costin-Gabriel Chiru, Stefan Trausan-Matu

专题命中 跨模态检索 :multimodal(abstract);分类 cs.CL

Comments Accepted at SemEval-2020, 7 pages, 1 figure, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.06351 2020-11-02 cs.CL 57%

CAPT: Contrastive Pre-Training for Learning Denoised Sequence Representations

Fuli Luo, Pengcheng Yang, Shicheng Li, Xuancheng Ren, Xu Sun

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CL

Comments Corrected typos

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.13668 2020-10-27 cs.LG cs.AI 57%

GraphMDN: Leveraging graph structure and deep learning to solve inverse problems

Tuomas P. Oikarinen, Daniel C. Hannah, Sohrob Kazerounian

专题命中 跨模态检索 :multi-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.04858 2020-08-31 cs.CV 57%

KBGN: Knowledge-Bridge Graph Network for Adaptive Vision-Text Reasoning in Visual Dialogue

Xiaoze Jiang, Siyi Du, Zengchang Qin, Yajing Sun, Jing Yu

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments Accepted by the 28th ACM International Conference on Multimedia (ACM MM 2020), Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.09609 2020-08-25 cs.CV 57%

Symbiotic Adversarial Learning for Attribute-based Person Search

Yu-Tong Cao, Jingya Wang, Dacheng Tao

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments 17 pages, 5 figures. Accepted to ECCV2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.06551 2020-08-18 cs.CV cs.IT cs.LG math.IT 57%

Sketch-Guided Object Localization in Natural Images

Aditay Tripathi, Rajath R Dani, Anand Mishra, Anirban Chakraborty

专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV

Comments ECCV 2020 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏