arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 4882 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 其他多模态 4882 篇

2304.09421 2023-04-20 cs.CL cs.CV cs.LG cs.SI 62%

TieFake: Title-Text Similarity and Emotion-Aware Fake News Detection

Quanjiang Guo, Zhao Kang, Ling Tian, Zhouguo Chen

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.CL

Comments Appear on IJCNN 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.01258 2023-03-03 cs.CL cs.AI cs.LG 62%

Domain-adapted large language models for classifying nuclear medicine reports

Zachary Huemann, Changhee Lee, Junjie Hu, Steve Y. Cho, Tyler Bradshaw

专题命中 其他多模态 :multimodal(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.13631 2023-02-28 eess.IV cs.AI cs.CV cs.LG q-bio.QM 62%

Curriculum Based Multi-Task Learning for Parkinson's Disease Detection

Nikhil J. Dhinagar, Conor Owens-Walton, Emily Laltoo, Christina P. Boyle, Yao-Liang Chen, Philip Cook, Corey McMillan, Chih-Chien Tsai, J-J Wang, Yih-Ru Wu, Ysbrand van der Werf, Paul M. Thompson

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments Accepted for publication at the 20th IEEE International Symposium on Biomedical Imaging, ISBI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.12050 2023-02-24 cs.CL cs.AI cs.LG cs.LO 62%

SPINDLE: Spinning Raw Text into Lambda Terms with Graph Attention

Konstantinos Kogkalidis, Michael Moortgat, Richard Moot

专题命中 其他多模态 :multimodal(abstract);分类 cs.CL、cs.AI

Comments EACL23 System Demonstrations

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.02979 2023-02-07 cs.CV cs.AI cs.LG 62%

Learning disentangled representations for explainable chest X-ray classification using Dirichlet VAEs

Rachael Harkness, Alejandro F Frangi, Kieran Zucker, Nishant Ravikumar

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.AI

Comments 13 pages, 8 figures, to be published in SPIE Medical Imaging 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.02940 2023-02-07 cs.CV cs.AI cs.HC 62%

Integrating Eye-Gaze Data into CXR DL Approaches: A Preliminary study

André Luís, Chihcheng Hsieh, Isabel Blanco Nobre, Sandra Costa Sousa, Anderson Maciel, Catarina Moreira, Joaquim Jorge

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments A version of this paper has been accepted for presentation at the 2nd XR Health workshop - XR Technologies for Healthcare and Wellbeing https://ieeevr.org/2023/contribute/workshoppapers/#XRHealth

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.04272 2022-12-09 cs.LG cs.AI cs.CL cs.SI 62%

A Modality-level Explainable Framework for Misinformation Checking in Social Networks

Vítor Lourenço, Aline Paes

专题命中 其他多模态 :multimodal(abstract);分类 cs.CL、cs.AI

Comments Accepted to publication at LatinX in AI workshop at the Thirty-sixth Conference on Neural Information Processing Systems, LXAI @ NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.14544 2022-12-02 cs.CV cs.AI cs.LG 62%

Target-Free Text-guided Image Manipulation

Wan-Cyuan Fan, Cheng-Fu Yang, Chiao-An Yang, Yu-Chiang Frank Wang

专题命中 其他多模态 :cross-modal(abstract);分类 cs.CV、cs.AI

Comments AAAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.15559 2022-10-28 cs.CV cs.AI cs.LG cs.RO eess.IV 62%

Robust Monocular Localization of Drones by Adapting Domain Maps to Depth Prediction Inaccuracies

Priyesh Shukla, Sureshkumar S., Alex C. Stutts, Sathya Ravi, Theja Tulabandhula, Amit R. Trivedi

专题命中 其他多模态 :cross-modal(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.05815 2022-10-13 cs.CV cs.CL 62%

Underspecification in Scene Description-to-Depiction Tasks

Ben Hutchinson, Jason Baldridge, Vinodkumar Prabhakaran

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.02628 2022-10-11 cs.IR cs.AI cs.CL 62%

HYCEDIS: HYbrid Confidence Engine for Deep Document Intelligence System

Bao-Sinh Nguyen, Quang-Bach Tran, Tuan-Anh Nguyen Dang, Duc Nguyen, Hung Le

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CL、cs.AI

Comments Document Intelligence @ KDD 2021 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.12505 2022-09-19 cs.CL cs.CV 62%

AiM: Taking Answers in Mind to Correct Chinese Cloze Tests in Educational Applications

Yusen Zhang, Zhongli Li, Qingyu Zhou, Ziyi Liu, Chao Li, Mina Ma, Yunbo Cao, Hongzhi Liu

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.CL

Comments Accepted to COLING 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.00448 2022-09-02 cs.AI cs.CL 62%

Intelligent Traffic Monitoring with Hybrid AI

Ehsan Qasemi, Alessandro Oltramari

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CL、cs.AI

Comments IJCAI Workshop on Artificial Intelligence for Autonomous Driving (AI4AD) 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.10414 2022-08-23 cs.CV cs.AI cs.HC 62%

MetaFi: Device-Free Pose Estimation via Commodity WiFi for Metaverse Avatar Simulation

Jianfei Yang, Yunjiao Zhou, He Huang, Han Zou, Lihua Xie

专题命中 其他多模态 :cross-modal(abstract);分类 cs.CV、cs.AI

Comments 6 pages, 3 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.11623 2022-07-26 cs.HC cs.AI cs.CV cs.LG 62%

A Simplistic and Cost-Effective Design for Real-World Development of an Ambient Assisted Living System for Fall Detection and Indoor Localization: Proof of Concept

Nirmalya Thakur, Chia Y. Han

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.09717 2022-07-18 cs.CV cs.MM 62%

EKTVQA: Generalized use of External Knowledge to empower Scene Text in Text-VQA

Arka Ujjal Dey, Ernest Valveny, Gaurav Harit

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.MM

Comments Accepted at IEEE Access

Journal ref IEEE.ACCESS 10 (2022) 72092-72106

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.05261 2022-07-13 cs.CL cs.AI cs.LG 62%

Building Korean Sign Language Augmentation (KoSLA) Corpus with Data Augmentation Technique

Changnam An, Eunkyung Han, Dongmyeong Noh, Ohkyoon Kwon, Sumi Lee, Hyunshim Han

专题命中 其他多模态 :multimodal(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.02639 2022-07-07 cs.CV cs.MM 62%

Adversarial Robustness of Visual Dialog

Lu Yu, Verena Rieser

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.12739 2022-06-24 cs.CV cs.CL cs.LG 62%

Modulating Bottom-Up and Top-Down Visual Processing via Language-Conditional Filters

İlker Kesen, Ozan Arkan Can, Erkut Erdem, Aykut Erdem, Deniz Yuret

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.CL

Comments 13 pages, 6 figures, 6 tables. Appeared in MULA Workshop at CVPR 2022

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, 2022, pp. 4610-4620

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.11224 2022-06-14 cs.CL cs.AI 62%

Introducing the diagrammatic semiotic mode

Tuomo Hiippala, John A. Bateman

专题命中 其他多模态 :multimodal(abstract);分类 cs.CL、cs.AI

Comments 16 pages; accepted at Diagrams 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.02714 2022-06-07 cs.CV cs.AI cs.LG 62%

FuSS: Fusing Superpixels for Improved Segmentation Consistency

Ian Nunes, Matheus B. Pereira, Hugo Oliveira, Jefersson A. Dos Santos, Marcus Poggi

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments submitted to IEEEACCESS. 19 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.14896 2022-03-29 cs.CV cs.AI cs.LG 62%

Multi-Task Learning for Visual Scene Understanding

Simon Vandenhende

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.AI

Comments PhD Thesis

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.14196 2022-03-29 cs.CV cs.AI 62%

HINT: Hierarchical Neuron Concept Explainer

Andong Wang, Wei-Ning Lee, Xiaojuan Qi

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments Accepted by CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.10001 2021-12-21 eess.IV cs.AI cs.CV cs.LG 62%

Cross-Domain Federated Learning in Medical Imaging

Vishwa S Parekh, Shuhao Lai, Vladimir Braverman, Jeff Leal, Steven Rowe, Jay J Pillai, Michael A Jacobs

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.AI

Comments Under Review for MIDL 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.09069 2021-12-17 cs.CV cs.AI cs.HC cs.LG eess.SP 62%

Progressive Graph Convolution Network for EEG Emotion Recognition

Yijin Zhou, Fu Li, Yang Li, Youshuo Ji, Guangming Shi, Wenming Zheng, Lijian Zhang, Yuanfang Chen, Rui Cheng

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.AI

Comments 11 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.03727 2021-12-08 cs.MM cs.AI 62%

RFGAN: RF-Based Human Synthesis

Cong Yu, Zhi Wu, Dongheng Zhang, Zhi Lu, Yang Hu, Yan Chen

专题命中 其他多模态 :cross-modal(abstract);分类 cs.AI、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.14339 2021-11-30 cs.CV cs.AI cs.NE 62%

Heterogeneous Visible-Thermal and Visible-Infrared Face Recognition using Unit-Class Loss and Cross-Modality Discriminator

Usman Cheema, Mobeen Ahmad, Dongil Han, Seungbin Moon

专题命中 其他多模态 :cross-modal(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.13309 2021-11-29 cs.CV cs.AI 62%

Data Augmented 3D Semantic Scene Completion with 2D Segmentation Priors

Aloisio Dourado, Frederico Guth, Teofilo de Campos

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.05191 2021-11-10 cs.CV cs.AI cs.LG eess.IV 62%

Does Thermal data make the detection systems more reliable?

Shruthi Gowda, Bahram Zonooz, Elahe Arani

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments Accepted at NeurIPS 2021 - ML4AD workshop (The code for this research is available at: https://github.com/NeurAI-Lab/MMC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.06486 2021-10-14 cs.CV cs.CL 62%

Understanding of Emotion Perception from Art

Digbalay Bose, Krishna Somandepalli, Souvik Kundu, Rimita Lahiri, Jonathan Gratch, Shrikanth Narayanan

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.CL

Comments 5 pages, 5 figures. Accepted at ICCV2021: 4th Workshop on Closing the loop between Vision and Language

详情

展开后加载摘要…

URL PDF HTML 收藏