arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 4882 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 其他多模态 4882 篇

2110.00423 2021-10-04 cs.CL cs.AI 62%

A Web Scale Entity Extraction System

Xuanting Cai, Quanbin Ma, Pan Li, Jianyu Liu, Qi Zeng, Zhengkan Yang, Pushkar Tripathi

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.09598 2021-08-26 cs.LG cs.AI cs.CV cs.NE 62%

SERF: Towards better training of deep neural networks using log-Softplus ERror activation Function

Sayan Nag, Mayukh Bhattacharyya

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.06207 2021-08-16 cs.IR cs.CL cs.MM 62%

Disentangling Hate in Online Memes

Rui Cao, Ziqing Fan, Roy Ka-Wei Lee, Wen-Haw Chong, Jing Jiang

专题命中 其他多模态 :multimodal(abstract);分类 cs.CL、cs.MM

Comments Paper accepted in ACM Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.05176 2021-07-13 cs.CV cs.CL 62%

Zero-Shot Compositional Concept Learning

Guangyue Xu, Parisa Kordjamshidi, Joyce Y. Chai

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.05556 2021-07-09 cs.CL cs.CV 62%

Sparse and Structured Visual Attention

Pedro Henrique Martins, Vlad Niculae, Zita Marinho, André Martins

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.00250 2021-06-29 cs.CL cs.CV 62%

ViTA: Visual-Linguistic Translation by Aligning Object Tags

Kshitij Gupta, Devansh Gautam, Radhika Mamidi

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.CL

Comments 7 pages, accepted at WAT-2021 co-located with ACL-IJCNLP 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.00442 2021-06-29 cs.LG cs.AI cs.CV cs.RO 62%

Touch-based Curiosity for Sparse-Reward Tasks

Sai Rajeswar, Cyril Ibrahim, Nitin Surya, Florian Golemo, David Vazquez, Aaron Courville, Pedro O. Pinheiro

专题命中 其他多模态 :cross-modal(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.07393 2021-06-15 cs.CL cs.AI cs.LG 62%

Grounding Language to Entities and Dynamics for Generalization in Reinforcement Learning

Austin W. Hanjie, Victor Zhong, Karthik Narasimhan

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CL、cs.AI

Comments Accepted to ICML 2021. Note author list and name changes from previous version

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.06008 2021-04-14 cs.CV cs.CL 62%

Disentangled Motif-aware Graph Learning for Phrase Grounding

Zongshen Mu, Siliang Tang, Jie Tan, Qiang Yu, Yueting Zhuang

专题命中 其他多模态 :cross-modal(abstract);分类 cs.CV、cs.CL

Comments 10 pages, 6 figures, AAAI 2021 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.09720 2021-04-01 cs.CV cs.AI 62%

Few-Shot Visual Grounding for Natural Human-Robot Interaction

Giorgos Tziafas, Hamidreza Kasaei

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.AI

Comments 6 pages, 4 figures, ICARSC2021 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.16510 2021-03-31 cs.HC cs.CV cs.GR cs.MM 62%

HapTable: An Interactive Tabletop Providing Online Haptic Feedback for Touch Gestures

Senem Ezgi Emgin, Amirreza Aghakhani, T. Metin Sezgin, Cagatay Basdogan

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.MM

Journal ref IEEE Transactions on Visualization and Computer Graphics, 2019, Vol. 25, No. 9, pp. 2749-2762

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.09602 2021-03-18 cs.SI cs.CL cs.CV 62%

On the Role of Images for Analyzing Claims in Social Media

Gullal S. Cheema, Sherzod Hakimov, Eric Müller-Budack, Ralph Ewerth

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.CL

Comments CLEOPATRA-2021 Workshop co-located with The Web Conf 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.12627 2021-01-01 cs.CL cs.AI cs.DB cs.LG 62%

Bridging Textual and Tabular Data for Cross-Domain Text-to-SQL Semantic Parsing

Xi Victoria Lin, Richard Socher, Caiming Xiong

专题命中 其他多模态 :cross-modal(abstract);分类 cs.CL、cs.AI

Comments EMNLP Findings 2020 long paper extended; 23 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.13615 2020-11-30 cs.CV cs.AI 62%

Manipulating Medical Image Translation with Manifold Disentanglement

Siyu Liu, Jason A. Dowling, Craig Engstrom, Peter B. Greer, Stuart Crozier, Shekhar S. Chandra

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.02959 2020-10-08 cs.CV cs.MM 62%

Using Sentences as Semantic Representations in Large Scale Zero-Shot Learning

Yannick Le Cacheux, Hervé Le Borgne, Michel Crucianu

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.00819 2020-03-17 cs.CL cs.AI cs.RO 62%

EDA: Enriching Emotional Dialogue Acts using an Ensemble of Neural Annotators

Chandrakant Bothe, Cornelius Weber, Sven Magg, Stefan Wermter

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CL、cs.AI

Comments Proceeding of the LREC 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.07809 2020-03-02 cs.LG cs.CL cs.CV cs.HC stat.ML 62%

Found in Translation: Learning Robust Joint Representations by Cyclic Translations Between Modalities

Hai Pham, Paul Pu Liang, Thomas Manzini, Louis-Philippe Morency, Barnabas Poczos

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.CL

Comments AAAI 2019, code available at https://github.com/hainow/MCTN

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.03481 2020-02-11 cs.RO cs.AI cs.CL cs.LG 62%

Improved and Scalable Online Learning of Spatial Concepts and Language Models with Mapping

Akira Taniguchi, Yoshinobu Hagiwara, Tadahiro Taniguchi, Tetsunari Inamura

专题命中 其他多模态 :multimodal(abstract);分类 cs.CL、cs.AI

Comments Accepted to Autonomous Robots (24 January 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.11106 2019-10-21 cs.CV cs.AI cs.LG 62%

Operational Neural Networks

Serkan Kiranyaz, Turker Ince, Alexandros Iosifidis, Moncef Gabbouj

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.AI

Comments 21 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.09822 2019-09-24 cs.CV cs.MM 62%

CANZSL: Cycle-Consistent Adversarial Networks for Zero-Shot Learning from Natural Language

Zhi Chen, Jingjing Li, Yadan Luo, Zi Huang, Yang Yang

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.MM

Comments WACV 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.06499 2019-05-17 cs.CV cs.AI 62%

Bimodal Stereo: Joint Shape and Pose Estimation from Color-Depth Image Pair

Chi Zhang, Yuehu Liu, Ying Wu, Qilin Zhang, Le Wang

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.AI

Comments Preprinted version on May 15, 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.05961 2019-05-16 cs.CY cs.CL cs.CV cs.LG 62%

Demographic Inference and Representative Population Estimates from Multilingual Social Media Data

Zijian Wang, Scott A. Hale, David Adelani, Przemyslaw A. Grabowicz, Timo Hartmann, Fabian Flöck, David Jurgens

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.CL

Comments 12 pages, 10 figures, Proceedings of the 2019 World Wide Web Conference (WWW '19)

Journal ref Proceedings of the 2019 World Wide Web Conference (WWW '19), May 13--17, 2019, San Francisco, CA, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.07708 2019-01-23 cs.AI cs.CL 62%

Knowledge will Propel Machine Understanding of Content: Extrapolating from Current Examples

Amit Sheth, Sujan Perera, Sanjaya Wijeratne

专题命中 其他多模态 :multimodal(abstract);分类 cs.CL、cs.AI

Comments There is a new version of this paper with new authors uploaded as arXiv:1707.05308, so this is an invalid entry

详情

展开后加载摘要…

URL PDF HTML 收藏
1702.08626 2018-06-19 cs.RO cs.AI cs.CV stat.ML 62%

Show, Attend and Interact: Perceivable Human-Robot Social Interaction through Neural Attention Q-Network

Ahmed Hussain Qureshi, Yutaka Nakamura, Yuichiro Yoshikawa, Hiroshi Ishiguro

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments 7 pages, 5 figures, accepted by IEEE-RAS ICRA'17

详情

展开后加载摘要…

URL PDF HTML 收藏
1802.00923 2018-02-06 cs.AI cs.CL cs.LG 62%

Multi-attention Recurrent Network for Human Communication Comprehension

Amir Zadeh, Paul Pu Liang, Soujanya Poria, Prateek Vij, Erik Cambria, Louis-Philippe Morency

专题命中 其他多模态 :multimodal(abstract);分类 cs.CL、cs.AI

Comments AAAI 2018 Oral Presentation

详情

展开后加载摘要…

URL PDF HTML 收藏
1801.06349 2018-01-22 cs.HC cs.AI cs.CV 62%

Proceedings of eNTERFACE 2015 Workshop on Intelligent Interfaces

Matei Mancas, Christian Frisson, Joëlle Tilmanne, Nicolas d'Alessandro, Petr Barborka, Furkan Bayansar, Francisco Bernard, Rebecca Fiebrink, Alexis Heloir, Edgar Hemery, Sohaib Laraba, Alexis Moinet, Fabrizio Nunnari, Thierry Ravet, Loïc Reboursière, Alvaro Sarasua, Mickaël Tits, Noé Tits, François Zajéga, Paolo Alborno, Ksenia Kolykhalova, Emma Frid, Damiano Malafronte, Lisanne Huis in't Veld, Hüseyin Cakmak, Kevin El Haddad, Nicolas Riche, Julien Leroy, Pierre Marighetto, Bekir Berker Türker, Hossein Khaki, Roberto Pulisci, Emer Gilmartin, Fasih Haider, Kübra Cengiz, Martin Sulir, Ilaria Torre, Shabbir Marzban, Ramazan Yazıcı, Furkan Burak Bâgcı, Vedat Gazi Kılı, Hilal Sezer, Sena Büsra Yenge, Charles-Alexandre Delestage, Sylvie Leleu-Merviel, Muriel Meyer-Chemenska, Daniel Schmitt, Willy Yvart, Stéphane Dupont, Ozan Can Altiok, Aysegül Bumin, Ceren Dikmen, Ivan Giangreco, Silvan Heller, Emre Külah, Gueorgui Pironkov, Luca Rossetto, Yusuf Sahillioglu, Heiko Schuldt, Omar Seddati, Yusuf Setinkaya, Metin Sezgin, Claudiu Tanase, Emre Toyan, Sean Wood, Doguhan Yeke, Françcois Rocca, Pierre-Henri De Deken, Alessandra Bandrabur, Fabien Grisard, Axel Jean-Caurant, Vincent Courboulay, Radhwan Ben Madhkour, Ambroise Moreau

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments 159 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1712.00733 2017-12-05 cs.CV cs.CL 62%

Incorporating External Knowledge to Answer Open-Domain Visual Questions with Dynamic Memory Networks

Guohao Li, Hang Su, Wenwu Zhu

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CV、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.05143 2017-06-19 cs.SI cs.AI cs.CY cs.MM 62%

AI-Powered Social Bots

Terrence Adams

专题命中 其他多模态 :multi-modal(abstract);分类 cs.AI、cs.MM

Comments 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1607.02660 2016-07-12 cs.HC cs.AI cs.CV 62%

Augmenting Supervised Emotion Recognition with Rule-Based Decision Model

Amol Patwardhan, Gerald Knapp

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments 8 pages, 6 figures, 23 tables, IEEE TAC (in review)

详情

展开后加载摘要…

URL PDF HTML 收藏
1606.03333 2016-06-13 cs.MM cs.CL cs.IR 62%

Automatic Genre and Show Identification of Broadcast Media

Mortaza Doulaty, Oscar Saz, Raymond W. M. Ng, Thomas Hain

专题命中 其他多模态 :multi-modal(abstract);分类 cs.CL、cs.MM

Comments Proc. of 17th Interspeech (2016), San Francisco, California, USA

详情

展开后加载摘要…

URL PDF HTML 收藏