arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 6897 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态训练与对齐 6897 篇

1910.04066 2019-10-10 cs.CV 79%

Deep Convolutional Neural Network for Multi-modal Image Restoration and Fusion

Xin Deng, Pier Luigi Dragotti

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.11464 2019-09-26 cs.CV 79%

Multi-modal segmentation with missing MR sequences using pre-trained fusion networks

Karin van Garderen, Marion Smits, Stefan Klein

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted at MICCAI MIL3ID workshop 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.07846 2019-09-18 cs.CV cs.LG 79%

Multimodal Multitask Representation Learning for Pathology Biobank Metadata Prediction

Wei-Hung Weng, Yuannan Cai, Angela Lin, Fraser Tan, Po-Hsuan Cameron Chen

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments preprint version

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.00420 2019-09-18 cs.LG cs.CV stat.ML 79%

Multi-Label Product Categorization Using Multi-Modal Fusion Models

Pasawee Wirojwatanakul, Artit Wangperawong

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.02031 2019-08-23 eess.IV cs.CV 79%

OctopusNet: A Deep Learning Segmentation Network for Multi-modal Medical Images

Yu Chen, Jiawei Chen, Dong Wei, Yuexiang Li, Yefeng Zheng

专题命中 多模态训练与对齐 :multi-modal(title);cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.10081 2019-07-25 cs.CV 79%

Multimodal Age and Gender Classification Using Ear and Profile Face Images

Dogucan Yaman, Fevziye Irem Eyiokur, Hazım Kemal Ekenel

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments 8 pages, 4 figures, accepted for CVPR 2019 - Workshop on Biometrics

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.00295 2019-06-04 cs.CL 79%

Multimodal Transformer for Unaligned Multimodal Language Sequences

Yao-Hung Hubert Tsai, Shaojie Bai, Paul Pu Liang, J. Zico Kolter, Louis-Philippe Morency, Ruslan Salakhutdinov

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.03619 2019-05-17 cs.CV 79%

Beyond Bilinear: Generalized Multimodal Factorized High-order Pooling for Visual Question Answering

Zhou Yu, Jun Yu, Chenchao Xiang, Jianping Fan, Dacheng Tao

专题命中 多模态训练与对齐 :multimodal(title);multi-modal(abstract);分类 cs.CV

Comments 13 pages, 9 figures. arXiv admin note: substantial text overlap with arXiv:1708.01471

Journal ref IEEE Transactions On Neural Networks And Learning Systems, Vol. 26, No. 10, October 2015, Pp. 2275-2290

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.08798 2019-05-01 cs.CV 79%

A scene perception system for visually impaired based on object detection and classification using multi-modal DCNN

Baljit Kaur, Jhilik Bhattacharya

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments 33pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.10450 2019-04-24 cs.LG cs.SD eess.AS stat.ML 79%

Latent Variable Algorithms for Multimodal Learning and Sensor Fusion

Lijiang Guo

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 eess.AS

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.10191 2019-03-11 cs.RO cs.AI cs.LG 79%

Making Sense of Vision and Touch: Self-Supervised Learning of Multimodal Representations for Contact-Rich Tasks

Michelle A. Lee, Yuke Zhu, Krishnan Srinivasan, Parth Shah, Silvio Savarese, Li Fei-Fei, Animesh Garg, Jeannette Bohg

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.AI

Comments ICRA 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.02967 2019-03-05 cs.CV 79%

HyperDense-Net: A hyper-densely connected CNN for multi-modal image segmentation

Jose Dolz, Karthik Gopinath, Jing Yuan, Herve Lombaert, Christian Desrosiers, Ismail Ben Ayed

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Paper accepted at IEEE TMI in October 2018. Last version of this paper updates the reference to the IEEE TMI paper which compares the submissions to the iSEG 2017 MICCAI Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.03938 2019-02-12 cs.CV 79%

MISO: Mutual Information Loss with Stochastic Style Representations for Multimodal Image-to-Image Translation

Sanghyeon Na, Seungjoo Yoo, Jaegul Choo

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.12829 2018-12-26 cs.CV 79%

Cross-Modal Attentional Context Learning for RGB-D Object Detection

Guanbin Li, Yukang Gan, Hejun Wu, Nong Xiao, Liang Lin

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments Accept as a regular paper to IEEE Transactions on Image Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.09276 2018-12-24 cs.LG cs.CV stat.ML 79%

Multimodal Sensor Fusion In Single Thermal image Super-Resolution

Feras Almasri, Olivier Debeir

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.08615 2018-11-28 cs.LG cs.CL 79%

Unsupervised Multimodal Representation Learning across Medical Images and Reports

Tzu-Ming Harry Hsu, Wei-Hung Weng, Willie Boag, Matthew McDermott, Peter Szolovits

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments Machine Learning for Health (ML4H) Workshop at NeurIPS 2018 arXiv:1811.07216

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.08305 2018-11-21 cs.CV 79%

IVD-Net: Intervertebral disc localization and segmentation in MRI with a multi-modal UNet

Jose Dolz, Christian Desrosiers, Ismail Ben Ayed

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Manuscript submitted to the Proceedings of the MICCAI 2018 IVD Challenge. arXiv admin note: text overlap with arXiv:1810.07003

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.04697 2018-11-13 cs.CL 79%

CUNI System for the WMT18 Multimodal Translation Task

Jindřich Helcl, Jindřich Libovický, Dušan Variš

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments Published at WMT18

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.02745 2018-11-08 cs.CV 79%

Y^2Seq2Seq: Cross-Modal Representation Learning for 3D Shape and Text by Joint Reconstruction and Prediction of View and Word Sequences

Zhizhong Han, Mingyang Shang, Xiyang Wang, Yu-Shen Liu, Matthias Zwicker

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments To be pubilished at AAAI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.06233 2018-11-05 cs.CV 79%

Robust Deep Multi-modal Learning Based on Gated Information Fusion Network

Jaekyum Kim, Junho Koh, Yecheol Kim, Jaehyung Choi, Youngbae Hwang, Jun Won Choi

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments 2018 Asian Conference on Computer Vision (ACCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.01220 2018-10-22 cs.CV 79%

Multi-Modal Multi-Scale Deep Learning for Large-Scale Image Annotation

Yulei Niu, Zhiwu Lu, Ji-Rong Wen, Tao Xiang, Shih-Fu Chang

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Submited to IEEE TIP

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.02001 2018-10-05 cs.CV 79%

Image and Encoded Text Fusion for Multi-Modal Classification

Ignazio Gallo, Alessandro Calefati, Shah Nawaz, Muhammad Kamran Janjua

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted to DICTA 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.08993 2018-09-28 cs.CV 79%

Improved Semantic Stixels via Multimodal Sensor Fusion

Florian Piewak, Peter Pinggera, Markus Enzweiler, David Pfeiffer, Marius Zöllner

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.04732 2018-08-16 cs.CV cs.LG stat.ML 79%

Multimodal Unsupervised Image-to-Image Translation

Xun Huang, Ming-Yu Liu, Serge Belongie, Jan Kautz

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments Accepted by ECCV 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.04836 2018-07-17 cs.CV 79%

Disjoint Mapping Network for Cross-modal Matching of Voices and Faces

Yandong Wen, Mahmoud Al Ismail, Weiyang Liu, Bhiksha Raj, Rita Singh

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments Tech report

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.03232 2018-07-10 eess.SP cs.CV physics.med-ph 79%

Robust Heartbeat Detection from Multimodal Data via CNN-based Generalizable Information Fusion

B S Chandra, C S Sastry, S Jana

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.00864 2018-07-04 cs.CV 79%

Semi-supervised Learning: Fusion of Self-supervised, Supervised Learning, and Multimodal Cues for Tactical Driver Behavior Detection

Athma Narayanan, Yi-Ting Chen, Srikanth Malla

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.00064 2018-06-04 cs.AI cs.LG stat.ML 79%

Efficient Low-rank Multimodal Fusion with Modality-Specific Factors

Zhun Liu, Ying Shen, Varun Bharadhwaj Lakshminarasimhan, Paul Pu Liang, Amir Zadeh, Louis-Philippe Morency

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.AI

Comments * Equal contribution. 10 pages. Accepted by ACL 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.08660 2018-05-23 cs.CL 79%

Multimodal Affective Analysis Using Hierarchical Attention Strategy with Word-Level Alignment

Yue Gu, Kangning Yang, Shiyu Fu, Shuhong Chen, Xinyu Li, Ivan Marsic

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments Accepted by ACL 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1801.00532 2018-01-03 cs.CL 79%

Learning Multimodal Word Representation via Dynamic Fusion Methods

Shaonan Wang, Jiajun Zhang, Chengqing Zong

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments To be appear in AAAI-18

详情

展开后加载摘要…

URL PDF HTML 收藏