arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 3450 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 跨模态检索 3450 篇

2306.15796 2023-06-29 cs.AI 79%

ConKI: Contrastive Knowledge Injection for Multimodal Sentiment Analysis

Yakun Yu, Mingjun Zhao, Shi-ang Qi, Feiran Sun, Baoxun Wang, Weidong Guo, Xiaoli Wang, Lei Yang, Di Niu

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.AI

Comments Accepted by ACL Findings 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.06066 2023-06-12 cs.CV cs.LG 79%

Multi-level Cross-modal Feature Alignment via Contrastive Learning towards Zero-shot Classification of Remote Sensing Image Scenes

Chun Liu, Suqiang Ma, Zheng Li, Wei Yang, Zhigang Han

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.04272 2023-06-08 cs.CV 79%

On the Generalization of Multi-modal Contrastive Learning

Qi Zhang, Yifei Wang, Yisen Wang

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.17663 2023-05-30 cs.CL 79%

Lexical Retrieval Hypothesis in Multimodal Context

Po-Ya Angela Wang, Pin-Er Chen, Hsin-Yu Chou, Yu-Hsiang Tseng, Shu-Kai Hsieh

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.16566 2023-05-29 cs.CV 79%

Integrating Listwise Ranking into Pairwise-based Image-Text Retrieval

Zheng Li, Caili Guo, Xin Wang, Zerun Feng, Yanjun Wang

专题命中 跨模态检索 :image-text(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.04239 2023-05-09 cs.CV cs.IR 79%

Instance-Variant Loss with Gaussian RBF Kernel for 3D Cross-modal Retriveal

Zhitao Liu, Zengyu Liu, Jiwei Wei, Guan Wang, Zhenjiang Du, Ning Xie, Heng Tao Shen

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.01915 2023-05-04 cs.IR cs.MM 79%

Denoising Multi-modal Sequential Recommenders with Contrastive Learning

Dong Yao, Shengyu Zhang, Zhou Zhao, Jieming Zhu, Wenqiao Zhang, Rui Zhang, Xiaofei He, Fei Wu

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.13357 2023-04-27 cs.CV cs.IR 79%

Deep Lifelong Cross-modal Hashing

Liming Xu, Hanqi Li, Bochuan Zheng, Weisheng Li, Jiancheng Lv

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.13103 2023-04-27 cs.CR cs.AI 79%

HyMo: Vulnerability Detection in Smart Contracts using a Novel Multi-Modal Hybrid Model

Mohammad Khodadadi, Jafar Tahmoresnezhad

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.04224 2023-04-11 cs.CV cs.LG 79%

Pix2Map: Cross-modal Retrieval for Inferring Street Maps from Images

Xindi Wu, KwunFung Lau, Francesco Ferroni, Aljoša Ošep, Deva Ramanan

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments 12 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.14080 2023-03-31 cs.CV 79%

Best of Both Worlds: Multimodal Contrastive Learning with Tabular and Imaging Data

Paul Hager, Martin J. Menten, Daniel Rueckert

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV

Comments Accepted in CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.04188 2023-03-28 cs.CV 79%

DepthFormer: Multimodal Positional Encodings and Cross-Input Attention for Transformer-Based Segmentation Networks

Francesco Barbato, Giulia Rizzoli, Pietro Zanuttigh

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV

Comments Accepted at ICASSP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.10839 2023-03-22 cs.CV 79%

MXM-CLR: A Unified Framework for Contrastive Learning of Multifold Cross-Modal Representations

Ye Wang, Bowei Jiang, Changqing Zou, Rui Ma

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments 16 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.10249 2023-03-21 eess.IV cs.CV 79%

MRIS: A Multi-modal Retrieval Approach for Image Synthesis on Diverse Modalities

Boqi Chen, Marc Niethammer

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.03597 2023-03-21 cs.CV 79%

Cross-Modality Sub-Image Retrieval using Contrastive Multimodal Image Representations

Eva Breznik, Elisabeth Wetzer, Joakim Lindblad, Nataša Sladoje

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.05692 2023-03-13 cs.CV 79%

Semantic-Preserving Augmentation for Robust Image-Text Retrieval

Sunwoo Kim, Kyuhong Shim, Luong Trung Nguyen, Byonghyo Shim

专题命中 跨模态检索 :image-text(title,abstract);分类 cs.CV

Comments Accepted to ICASSP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.11352 2023-02-23 cs.CV 79%

X-TRA: Improving Chest X-ray Tasks with Cross-Modal Retrieval Augmentation

Tom van Sonsbeek, Marcel Worring

专题命中 跨模态检索 :cross-modal(title);multi-modal(abstract);分类 cs.CV

Comments IPMI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.04754 2023-02-15 cs.CV cs.IR 79%

Improving Visual-Semantic Embeddings by Learning Semantically-Enhanced Hard Negatives for Cross-modal Information Retrieval

Yan Gong, Georgina Cosma

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Journal ref Pattern Recognition 137 (2023): 109272

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.04742 2023-01-13 cs.CV 79%

HADA: A Graph-based Amalgamation Framework in Image-text Retrieval

Manh-Duy Nguyen, Binh T. Nguyen, Cathal Gurrin

专题命中 跨模态检索 :image-text(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.04948 2022-12-19 cs.CV 79%

Transformer-based Cross-Modal Recipe Embeddings with Large Batch Training

Jing Yang, Junwen Chen, Keiji Yanai

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments Accepted at MMM2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.16208 2022-12-09 cs.CV 79%

SLAN: Self-Locator Aided Network for Cross-Modal Understanding

Jiang-Tian Zhai, Qi Zhang, Tong Wu, Xing-Yu Chen, Jiang-Jiang Liu, Bo Ren, Ming-Ming Cheng

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.01612 2022-12-06 cs.CL 79%

Named Entity and Relation Extraction with Multi-Modal Retrieval

Xinyu Wang, Jiong Cai, Yong Jiang, Pengjun Xie, Kewei Tu, Wei Lu

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CL

Comments Findings of EMNLP 2022. Code is publicly available at http://github.com/modelscope/adaseq/examples/MoRe

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.16262 2022-11-30 eess.IV cs.CV 79%

Is Image-to-Image Translation the Panacea for Multimodal Image Registration? A Comparative Study

Jiahao Lu, Johan Öfverstedt, Joakim Lindblad, Nataša Sladoje

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV

Comments 37 pages, 10 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.12032 2022-11-24 cs.CV 79%

PointCMC: Cross-Modal Multi-Scale Correspondences Learning for Point Cloud Understanding

Honggu Zhou, Xiaogang Peng, Jiawei Mao, Zizhao Wu, Ming Zeng

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments In order to revise the paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.11256 2022-11-22 cs.CL 79%

UniMSE: Towards Unified Multimodal Sentiment Analysis and Emotion Recognition

Guimin Hu, Ting-En Lin, Yi Zhao, Guangming Lu, Yuchuan Wu, Yongbin Li

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CL

Comments Accepted to EMNLP 2022 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.01884 2022-11-08 cs.LG cs.AI 79%

Graph Neural Networks for Multimodal Single-Cell Data Integration

Hongzhi Wen, Jiayuan Ding, Wei Jin, Yiqi Wang, Yuying Xie, Jiliang Tang

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.AI

Comments Accepted by KDD 2022 Applied Data Science Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.10486 2022-10-20 cs.CV cs.LG 79%

Cross-Modal Fusion Distillation for Fine-Grained Sketch-Based Image Retrieval

Abhra Chaudhuri, Massimiliano Mancini, Yanbei Chen, Zeynep Akata, Anjan Dutta

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments British Machine Vision Conference (BMVC) 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.08908 2022-10-18 cs.CV 79%

Cross-modal Semantic Enhanced Interaction for Image-Sentence Retrieval

Xuri Ge, Fuhai Chen, Songpei Xu, Fuxiang Tao, Joemon M. Jose

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments accepted to WACV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.04341 2022-10-11 cs.CV 79%

ConTra: (Con)text (Tra)nsformer for Cross-Modal Video Retrieval

Adriano Fragomeni, Michael Wray, Dima Damen

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments Accepted in ACCV 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.01416 2022-09-07 cs.AI cs.DB 79%

MMKGR: Multi-hop Multi-modal Knowledge Graph Reasoning

Shangfei Zheng, Weiqing Wang, Jianfeng Qu, Hongzhi Yin, Wei Chen, Lei Zhao

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏