arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 6887 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态训练与对齐 6887 篇

2310.16936 2023-10-31 cs.CV cs.LG 79%

Diagnosing Alzheimer's Disease using Early-Late Multimodal Data Fusion with Jacobian Maps

Yasmine Mustafa, Tie Luo

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments To be published in Proceedings of 2023 IEEE Healthcom, December 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.16856 2023-10-27 cs.CV 79%

GraFT: Gradual Fusion Transformer for Multimodal Re-Identification

Haoli Yin, Jiayao Li, Eva Schiller, Luke McDermott, Daniel Cummings

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments 3 Borderline Reviews at WACV, 8 pages, 5 figures, 8 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.14496 2023-10-24 cs.MM 79%

Redundancy-Adaptive Multimodal Learning for Imperfect Data

Mengxi Chen, Jiangchao Yao, Linyu Xing, Yu Wang, Ya Zhang, Yanfeng Wang

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.07871 2023-10-23 cs.AI 79%

Hierarchical Pretraining on Multimodal Electronic Health Records

Xiaochen Wang, Junyu Luo, Jiaqi Wang, Ziyi Yin, Suhan Cui, Yuan Zhong, Yaqing Wang, Fenglong Ma

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.AI

Comments Accepted by EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.12802 2023-10-23 cs.LG cs.AI q-bio.BM 79%

Otter-Knowledge: benchmarks of multimodal knowledge graph representation learning from different sources for drug discovery

Hoang Thanh Lam, Marco Luca Sbodio, Marcos Martínez Galindo, Mykhaylo Zayats, Raúl Fernández-Díaz, Víctor Valls, Gabriele Picco, Cesar Berrospi Ramis, Vanessa López

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.11910 2023-10-19 eess.IV cs.CV cs.LG 79%

Multi-modal Medical Neurological Image Fusion using Wavelet Pooled Edge Preserving Autoencoder

Manisha Das, Deep Gupta, Petia Radeva, Ashwini M Bakde

专题命中 多模态训练与对齐 :multi-modal(title);multimodal(abstract);分类 cs.CV

Comments 8 pages, 5 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.11896 2023-10-19 eess.IV cs.CV cs.LG 79%

A New Multimodal Medical Image Fusion based on Laplacian Autoencoder with Channel Attention

Payal Wankhede, Manisha Das, Deep Gupta, Petia Radeva, Ashwini M Bakde

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments 10 pages, 6 figures, % tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.10649 2023-10-17 cs.CV 79%

Cross-modal and Cross-domain Knowledge Transfer for Label-free 3D Segmentation

Jingyu Zhang, Huitong Yang, Dai-Jie Wu, Jacky Keung, Xuesong Li, Xinge Zhu, Yuexin Ma

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments 12 pages,4 figures,accepted

Journal ref Chinese Conference on Pattern Recognition and Computer Vision (PRCV) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.06365 2023-10-11 cs.CL 79%

Multi-Modal Knowledge Graph Transformer Framework for Multi-Modal Entity Alignment

Qian Li, Cheng Ji, Shu Guo, Zhaoji Liang, Lihong Wang, Jianxin Li

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03485 2023-10-10 eess.IV cs.CV cs.LG 79%

BTDNet: a Multi-Modal Approach for Brain Tumor Radiogenomic Classification

Dimitrios Kollias, Karanjot Vendal, Priyanka Gadhavi, Solomon Russom

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.08747 2023-10-06 cs.CV 79%

Unified Brain MR-Ultrasound Synthesis using Multi-Modal Hierarchical Representations

Reuben Dorent, Nazim Haouchine, Fryderyk Kögl, Samuel Joutard, Parikshit Juvekar, Erickson Torio, Alexandra Golby, Sebastien Ourselin, Sarah Frisken, Tom Vercauteren, Tina Kapur, William M. Wells

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted at MICCAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.02960 2023-10-05 cs.CV 79%

CoDA: Collaborative Novel Box Discovery and Cross-modal Alignment for Open-vocabulary 3D Object Detection

Yang Cao, Yihan Zeng, Hang Xu, Dan Xu

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments Accepted by NeurIPS 2023. Project Page: https://yangcaoai.github.io/publications/CoDA.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.02663 2023-10-05 cs.CV 79%

MedPrompt: Cross-Modal Prompting for Multi-Task Medical Image Translation

Xuhang Chen, Chi-Man Pun, Shuqiang Wang

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.01912 2023-10-04 eess.IV cs.CV cs.LG 79%

Improved Automatic Diabetic Retinopathy Severity Classification Using Deep Multimodal Fusion of UWF-CFP and OCTA Images

Mostafa El Habib Daho, Yihao Li, Rachid Zeghlache, Yapo Cedric Atse, Hugo Le Boité, Sophie Bonnin, Deborah Cosette, Pierre Deman, Laurent Borderie, Capucine Lepicard, Ramin Tadayoni, Béatrice Cochener, Pierre-Henri Conze, Mathieu Lamard, Gwenolé Quellec

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments Accepted preprint for presentation at MICCAI-OMIA 20023, Vancouver, Canada

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.06612 2023-09-29 cs.LG cs.CV 79%

Harmonic-NAS: Hardware-Aware Multimodal Neural Architecture Search on Resource-constrained Devices

Mohamed Imed Eddine Ghebriout, Halima Bouzidi, Smail Niar, Hamza Ouarnoughi

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments Accepted to the 15th Asian Conference on Machine Learning (ACML 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.15529 2023-09-28 eess.IV cs.CV cs.LG 79%

Missing-modality Enabled Multi-modal Fusion Architecture for Medical Data

Muyu Wang, Shiyu Fan, Yichen Li, Hui Chen

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.13650 2023-09-26 eess.AS cs.SD 79%

Cross-modal Alignment with Optimal Transport for CTC-based ASR

Xugang Lu, Peng Shen, Yu Tsao, Hisashi Kawai

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 eess.AS

Comments Accepted to IEEE ASRU 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.11811 2023-09-22 eess.SP cs.AI 79%

Multimodal Transformers for Wireless Communications: A Case Study in Beam Prediction

Yu Tian, Qiyang Zhao, Zine el abidine Kherroubi, Fouzi Boukhalfa, Kebin Wu, Faouzi Bader

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.09593 2023-09-19 cs.CV cs.IT cs.RO math.IT 79%

Mutual Information-calibrated Conformal Feature Fusion for Uncertainty-Aware Multimodal 3D Object Detection at the Edge

Alex C. Stutts, Danilo Erricolo, Sathya Ravi, Theja Tulabandhula, Amit Ranjan Trivedi

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05787 2023-09-13 cs.AI cs.HC cs.LG 79%

Adaptive User-centered Neuro-symbolic Learning for Multimodal Interaction with Autonomous Systems

Amr Gomaa, Michael Feld

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.AI

Comments AI&HCI Workshop accepted paper at ICML2023 and accepted at ICMI2023 Blue Sky Papers. arXiv admin note: text overlap with arXiv:2211.03539

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05608 2023-09-12 cs.CL cs.CE 79%

Incorporating Pre-trained Model Prompting in Multimodal Stock Volume Movement Prediction

Ruibo Chen, Zhiyuan Zhang, Yi Liu, Ruihan Bao, Keiko Harimoto, Xu Sun

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments 9 pages, 3 figures, 7 tables. Accepted by 2023 KDD Workshop on Machine Learning in Finance

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.07901 2023-09-12 cs.LG cs.CV 79%

Auxiliary Cross-Modal Representation Learning with Triplet Loss Functions for Online Handwriting Recognition

Felix Ott, David Rügamer, Lucas Heublein, Bernd Bischl, Christopher Mutschler

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Journal ref IEEE Access, volume 11, pages 94148-94172, August 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.04062 2023-09-11 cs.LG cs.AI physics.chem-ph 79%

3D Denoisers are Good 2D Teachers: Molecular Pretraining via Denoising and Cross-Modal Distillation

Sungjun Cho, Dae-Woong Jeong, Sung Moon Ko, Jinwoo Kim, Sehui Han, Seunghoon Hong, Honglak Lee, Moontae Lee

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.AI

Comments 16 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.00601 2023-09-08 cs.CV 79%

Multimodal Industrial Anomaly Detection via Hybrid Fusion

Yue Wang, Jinlong Peng, Jiangning Zhang, Ran Yi, Yabiao Wang, Chengjie Wang

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.02702 2023-09-07 cs.CV 79%

Gene-induced Multimodal Pre-training for Image-omic Classification

Ting Jin, Xingran Xie, Renjie Wan, Qingli Li, Yan Wang

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.01169 2023-09-06 cs.LG cs.AI 79%

End-to-End Learning on Multimodal Knowledge Graphs

W. X. Wilcke, P. Bloem, V. de Boer, R. H. van t Veer

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.AI

Comments Under submission. arXiv admin note: substantial text overlap with arXiv:2003.12383

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.09941 2023-09-04 cs.CV 79%

A Robust and Interpretable Deep Learning Framework for Multi-modal Registration via Keypoints

Alan Q. Wang, Evan M. Yu, Adrian V. Dalca, Mert R. Sabuncu

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted to Medical Image Analysis 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.14505 2023-09-01 eess.IV cs.CV cs.LG 79%

Transformer-based interpretable multi-modal data fusion for skin lesion classification

Theodor Cheslerean-Boghiu, Melia-Evelina Fleischmann, Theresa Willem, Tobias Lasser

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Submitted to IEEE JBHI in July 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.15846 2023-08-31 cs.CV 79%

Exploring Multi-Modal Contextual Knowledge for Open-Vocabulary Object Detection

Yifan Xu, Mengdan Zhang, Xiaoshan Yang, Changsheng Xu

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.13883 2023-08-29 eess.IV cs.CV 79%

ReFuSeg: Regularized Multi-Modal Fusion for Precise Brain Tumour Segmentation

Aditya Kasliwal, Sankarshanaa Sagaram, Laven Srivastava, Pratinav Seth, Adil Khan

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted at 9th edition of the Brain Lesion (BrainLes) workshop, MICCAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏