arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 6897 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态训练与对齐 6897 篇

2008.01388 2020-08-05 cs.CV 79%

Unsupervised Cross-Modal Alignment for Multi-Person 3D Pose Estimation

Jogendra Nath Kundu, Ambareesh Revanur, Govind Vitthal Waghmare, Rahul Mysore Venkatesh, R. Venkatesh Babu

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments ECCV 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.09777 2020-07-21 cs.CV 79%

Deep Representation Learning For Multimodal Brain Networks

Wen Zhang, Liang Zhan, Paul Thompson, Yalin Wang

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments 11 pages, 3 figures, MICCAI 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.12209 2020-07-21 eess.IV cs.CV 79%

JSSR: A Joint Synthesis, Segmentation, and Registration System for 3D Multi-Modal Image Alignment of Large-scale Pathological CT Scans

Fengze Liu, Jinzheng Cai, Yuankai Huo, Chi-Tung Cheng, Ashwin Raju, Dakai Jin, Jing Xiao, Alan Yuille, Le Lu, ChienHung Liao, Adam P Harrison

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments accepted to ECCV 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.08742 2020-07-20 cs.CL 79%

A Novel Graph-based Multi-modal Fusion Encoder for Neural Machine Translation

Yongjing Yin, Fandong Meng, Jinsong Su, Chulun Zhou, Zhengyuan Yang, Jie Zhou, Jiebo Luo

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.04901 2020-07-10 cs.CV 79%

Cross-Modal Weighting Network for RGB-D Salient Object Detection

Gongyang Li, Zhi Liu, Linwei Ye, Yang Wang, Haibin Ling

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments Accepted in ECCV2020. Code: https://github.com/MathLee/CMWNet

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.08913 2020-07-01 cs.CV 79%

Seeing Through Fog Without Seeing Fog: Deep Multimodal Sensor Fusion in Unseen Adverse Weather

Mario Bijelic, Tobias Gruber, Fahim Mannan, Florian Kraus, Werner Ritter, Klaus Dietmayer, Felix Heide

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Journal ref The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.09454 2020-06-18 q-bio.NC cs.CV cs.LG eess.IV 79%

Interpretable multimodal fusion networks reveal mechanisms of brain cognition

Wenxing Hu, Xianghe Meng, Yuntong Bai, Aiying Zhang, Biao Cai, Gemeng Zhang, Tony W. Wilson, Julia M. Stephen, Vince D. Calhoun, Yu-Ping Wang

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.08159 2020-06-16 cs.CV 79%

Survey on Deep Multi-modal Data Analytics: Collaboration, Rivalry and Fusion

Yang Wang

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Appearing at ACM TOMM, 26 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.04392 2020-06-04 cs.CV 79%

Adaptive and Azimuth-Aware Fusion Network of Multimodal Local Features for 3D Object Detection

Yonglin Tian, Kunfeng Wang, Yuang Wang, Yulin Tian, Zilei Wang, Fei-Yue Wang

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments Accepted by Neurocomputing

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.01431 2020-06-03 cs.CV 79%

Distribution Aligned Multimodal and Multi-Domain Image Stylization

Minxuan Lin, Fan Tang, Weiming Dong, Xiao Li, Chongyang Ma, Changsheng Xu

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.10987 2020-05-25 cs.CV 79%

Investigating Vulnerability to Adversarial Examples on Multimodal Data Fusion in Deep Learning

Youngjoon Yu, Hong Joo Lee, Byeong Cheon Kim, Jung Uk Kim, Yong Man Ro

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.04294 2020-05-18 cs.CV 79%

Mix and match networks: cross-modal alignment for zero-pair image-to-image translation

Yaxing Wang, Luis Herranz, Joost van de Weijer

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments Accepted by IJCV

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.11744 2020-04-27 cs.CV 79%

PipeNet: Selective Modal Pipeline of Fusion Network for Multi-Modal Face Anti-Spoofing

Qing Yang, Xia Zhu, Jong-Kae Fwu, Yun Ye, Ganmei You, Yuan Zhu

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted to appear in CVPR2020 WMF

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.09321 2020-04-21 cs.CV eess.IV 79%

Combining multimodal information for Metal Artefact Reduction: An unsupervised deep learning framework

Marta B. M. Ranzini, Irme Groothuis, Kerstin Kläser, M. Jorge Cardoso, Johann Henckel, Sébastien Ourselin, Alister Hart, Marc Modat

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments Accepted at IEEE International Symposium on Biomedical Imaging (ISBI) 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.13910 2020-04-17 cs.CV 79%

Attention-based Multi-modal Fusion Network for Semantic Scene Completion

Siqi Li, Changqing Zou, Yipeng Li, Xibin Zhao, Yue Gao

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted by AAAI 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.07011 2020-04-16 cs.CV 79%

Code-Aligned Autoencoders for Unsupervised Change Detection in Multimodal Remote Sensing Images

Luigi T. Luppino, Mads A. Hansen, Michael Kampffmeyer, Filippo M. Bianchi, Gabriele Moser, Robert Jenssen, Stian N. Anfinsen

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.06229 2020-04-15 cs.LG cs.CV 79%

Imitation Learning for Fashion Style Based on Hierarchical Multimodal Representation

Shizhu Liu, Shanglin Yang, Hui Zhou

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.08670 2020-04-01 cs.CV cs.LG 79%

MMTM: Multimodal Transfer Module for CNN Fusion

Hamid Reza Vaezi Joze, Amirreza Shaban, Michael L. Iuzzolino, Kazuhito Koishida

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Journal ref The IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.06788 2020-03-24 cs.CV 79%

GMM-UNIT: Unsupervised Multi-Domain and Multi-Modal Image-to-Image Translation via Attribute Gaussian Mixture Modeling

Yahui Liu, Marco De Nadai, Jian Yao, Nicu Sebe, Bruno Lepri, Xavier Alameda-Pineda

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments 27 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.08073 2020-03-19 cs.CV 79%

Unsupervised Multi-Modal Image Registration via Geometry Preserving Image-to-Image Translation

Moab Arar, Yiftach Ginger, Dov Danon, Ilya Leizerson, Amit Bermano, Daniel Cohen-Or

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.05611 2020-03-05 cs.CV cs.LG cs.RO 79%

UNO: Uncertainty-aware Noisy-Or Multimodal Fusion for Unanticipated Input Degradation

Junjiao Tian, Wesley Cheung, Nathan Glaser, Yen-Cheng Liu, Zsolt Kira

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments IEEE International Conference on Robotics and Automation (ICRA), 2020. IROS Workshop on the Importance of Uncertainty in Deep Learning for Robotics, 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.12573 2020-03-02 cs.CV 79%

MANet: Multimodal Attention Network based Point- View fusion for 3D Shape Recognition

Yaxin Zhao, Jichao Jiao, Tangkun Zhang

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments 8 pages,6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.09708 2020-02-25 cs.CV 79%

Robust Multimodal Brain Tumor Segmentation via Feature Disentanglement and Gated Fusion

Cheng Chen, Qi Dou, Yueming Jin, Hao Chen, Jing Qin, Pheng-Ann Heng

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments MICCAI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.05000 2020-02-13 cs.CV eess.IV 79%

Hi-Net: Hybrid-fusion Network for Multi-modal MR Image Synthesis

Tao Zhou, Huazhu Fu, Geng Chen, Jianbing Shen, Ling Shao

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments has been accepted by IEEE TMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.06673 2020-01-22 cs.RO cs.CV 79%

A Transfer Learning Approach to Cross-Modal Object Recognition: From Visual Observation to Robotic Haptic Exploration

Pietro Falco, Shuang Lu, Ciro Natale, Salvatore Pirozzi, Dongheui Lee

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Journal ref IEEE Transactions on Robotics ( Volume: 35 , Issue: 4 , Aug. 2019 )

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.02340 2019-12-17 cs.CV 79%

Static and Dynamic Fusion for Multi-modal Cross-ethnicity Face Anti-spoofing

Ajian Liu, Zichang Tan, Xuan Li, Jun Wan, Sergio Escalera, Guodong Guo, Stan Z. Li

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments 10 pages, 9 figures, conference

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.00930 2019-11-22 cs.CV 79%

Multi-Resolution Multi-Modal Sensor Fusion For Remote Sensing Data With Label Uncertainty

Xiaoxiao Du, Alina Zare

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.08479 2019-11-21 cs.CV 79%

Modal-aware Features for Multimodal Hashing

Haien Zeng, Hanjiang Lai, Hanlu Chu, Yong Tang, Jian Yin

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.02645 2019-10-21 cs.CV 79%

Weakly Aligned Cross-Modal Learning for Multispectral Pedestrian Detection

Lu Zhang, Xiangyu Zhu, Xiangyu Chen, Xu Yang, Zhen Lei, Zhiyong Liu

专题命中 多模态训练与对齐 :cross-modal(title);multimodal(abstract);分类 cs.CV

Comments Accepted by ICCV2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1802.08568 2019-10-17 cs.CV 79%

Indic Handwritten Script Identification using Offline-Online Multimodal Deep Network

Ayan Kumar Bhunia, Subham Mukherjee, Aneeshan Sain, Ankan Kumar Bhunia, Partha Pratim Roy, Umapada Pal

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments Accepted in Information Fusion, Elsevier

详情

展开后加载摘要…

URL PDF HTML 收藏