arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

视觉大模型 / VLM

视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。

共收录 7387 信号源:cs.CV, cs.AI, cs.LG

1. 视觉定位与Grounding 7387 篇

2011.06850 2020-11-16 cs.CV cs.AI 62%

Transductive Zero-Shot Learning using Cross-Modal CycleGAN

Patrick Bordes, Eloi Zablocki, Benjamin Piwowarski, Patrick Gallinari

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.08792 2020-09-21 cs.CV cs.AI 62%

Commands 4 Autonomous Vehicles (C4AV) Workshop Summary

Thierry Deruyttere, Simon Vandenhende, Dusan Grujicic, Yu Liu, Luc Van Gool, Matthew Blaschko, Tinne Tuytelaars, Marie-Francine Moens

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.00283 2020-07-21 cs.CV cs.CL cs.LG 62%

Learning to Generate Grounded Visual Captions without Localization Supervision

Chih-Yao Ma, Yannis Kalantidis, Ghassan AlRegib, Peter Vajda, Marcus Rohrbach, Zsolt Kira

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Comments ECCV 2020. Code is available at https://github.com/chihyaoma/cyclical-visual-captioning

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.09904 2020-06-18 cs.IR cs.CV cs.LG 62%

Learning Colour Representations of Search Queries

Paridhi Maheshwari, Manoj Ghuhan, Vishwa Vinay

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Comments Accepted as a full paper at SIGIR 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.02174 2020-06-04 cs.CL cs.AI cs.LG 62%

CompGuessWhat?!: A Multi-task Evaluation Framework for Grounded Language Learning

Alessandro Suglia, Ioannis Konstas, Andrea Vanzo, Emanuele Bastianelli, Desmond Elliott, Stella Frank, Oliver Lemon

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments Accepted to the Annual Conference of the Association for Computational Linguistics (ACL) 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.04567 2020-05-26 cs.AI cs.CL cs.LG 62%

Ecological Semantics: Programming Environments for Situated Language Understanding

Ronen Tamari, Gabriel Stanovsky, Dafna Shahaf, Reut Tsarfaty

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments Camera ready for Bridging AI and Cognitive Science (BAICS) workshop at ICLR2020. For interactive demos, see https://eco-sem.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.12827 2020-05-07 cs.LG cs.CV cs.NE stat.ML 62%

Entity Abstraction in Visual Model-Based Reinforcement Learning

Rishi Veerapaneni, John D. Co-Reyes, Michael Chang, Michael Janner, Chelsea Finn, Jiajun Wu, Joshua B. Tenenbaum, Sergey Levine

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Comments Accepted at CoRL 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.11618 2020-03-27 cs.CV cs.AI cs.CL 62%

VIOLIN: A Large-Scale Dataset for Video-and-Language Inference

Jingzhou Liu, Wenhu Chen, Yu Cheng, Zhe Gan, Licheng Yu, Yiming Yang, Jingjing Liu

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments Accepted to CVPR2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.00401 2019-11-25 cs.AI cs.CL cs.CV 62%

Learning To Follow Directions in Street View

Karl Moritz Hermann, Mateusz Malinowski, Piotr Mirowski, Andras Banki-Horvath, Keith Anderson, Raia Hadsell

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Journal ref AAAI 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.06769 2019-11-06 cs.AI cs.LG cs.RO 62%

Continuous Relaxation of Symbolic Planner for One-Shot Imitation Learning

De-An Huang, Danfei Xu, Yuke Zhu, Animesh Garg, Silvio Savarese, Li Fei-Fei, Juan Carlos Niebles

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments IROS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.05448 2019-10-15 cs.NE cs.CV cs.LG stat.ML 62%

Neural Memory Plasticity for Anomaly Detection

Tharindu Fernando, Simon Denman, David Ahmedt-Aristizabal, Sridha Sridharan, Kristin Laurens, Patrick Johnston, Clinton Fookes

专题命中 视觉定位与Grounding :visual question answering(abstract);分类 cs.CV、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.10657 2019-09-12 cs.RO cs.CV cs.LG 62%

From explanation to synthesis: Compositional program induction for learning from demonstration

Michael Burke, Svetlin Penkov, Subramanian Ramamoorthy

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Journal ref Proceedings of Robotics: Science and Systems (2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.06556 2019-08-20 cs.CL cs.AI cs.LG 62%

Transfer in Deep Reinforcement Learning using Knowledge Graphs

Prithviraj Ammanabrolu, Mark O. Riedl

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.08313 2019-07-22 cs.AI cs.LG 62%

Learning High-Level Planning Symbols from Intrinsically Motivated Experience

Angelo Oddi, Riccardo Rasconi, Emilio Cartoni, Gabriele Sartor, Gianluca Baldassarre, Vieri Giuliano Santucci

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.02738 2019-06-10 cs.CL cs.AI cs.LG 62%

Conversing by Reading: Contentful Neural Conversation with On-demand Machine Reading

Lianhui Qin, Michel Galley, Chris Brockett, Xiaodong Liu, Xiang Gao, Bill Dolan, Yejin Choi, Jianfeng Gao

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments ACL 2019 long paper

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.05426 2019-04-12 cs.CL cs.AI cs.LG 62%

A Grounded Unsupervised Universal Part-of-Speech Tagger for Low-Resource Languages

Ronald Cardenas, Ying Lin, Heng Ji, Jonathan May

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments NAACL-HLT 2019, 12 pages, code available at https://github.com/isi-nlp/universal-cipher-pos-tagging

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.03885 2019-04-09 cs.CV cs.CL cs.LG 62%

Referring to Objects in Videos using Spatio-Temporal Identifying Descriptions

Peratham Wiriyathammabhum, Abhinav Shrivastava, Vlad I. Morariu, Larry S. Davis

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.09243 2019-03-25 cs.RO cs.AI cs.CL cs.LG 62%

Inferring Compact Representations for Efficient Natural Language Understanding of Robot Instructions

Siddharth Patki, Andrea F. Daniele, Matthew R. Walter, Thomas M. Howard

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments Accepted to ICRA 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1802.01433 2018-08-15 cs.CL cs.AI cs.LG 62%

Interactive Grounded Language Acquisition and Generalization in a 2D World

Haonan Yu, Haichao Zhang, Wei Xu

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments ICLR 2018 (Figure 6 caption improved)

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.01720 2018-04-09 cs.CV cs.CL cs.LG 62%

Finding beans in burgers: Deep semantic-visual embedding with localization

Martin Engilberge, Louis Chevallier, Patrick Pérez, Matthieu Cord

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Comments Accepted to CVPR2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.11017 2017-11-30 cs.AI cs.CL cs.CV cs.RO cs.SD eess.AS 62%

HoME: a Household Multimodal Environment

Simon Brodeur, Ethan Perez, Ankesh Anand, Florian Golemo, Luca Celotti, Florian Strub, Jean Rouat, Hugo Larochelle, Aaron Courville

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments Presented at NIPS 2017's Visually-Grounded Interaction and Language Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
1710.04076 2017-10-12 cs.RO cs.AI cs.CV 62%

Deep Semantic Abstractions of Everyday Human Activities: On Commonsense Representations of Human Interactions

Jakob Suchan, Mehul Bhatt

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments In ROBOT 2017: Third Iberian Robotics Conference. Escuela Técnica Superior de Ingeniería, Sevilla (Spain) (November 22-24, 2017). https://grvc.us.es/robot2017/ (to appear). arXiv admin note: substantial text overlap with arXiv:1709.05293

详情

展开后加载摘要…

URL PDF HTML 收藏
1605.09410 2017-07-14 cs.LG cs.CV 62%

End-to-End Instance Segmentation with Recurrent Attention

Mengye Ren, Richard S. Zemel

专题命中 视觉定位与Grounding :visual question answering(abstract);分类 cs.CV、cs.LG

Comments CVPR 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1701.08251 2017-04-21 cs.CL cs.AI cs.CV 62%

Image-Grounded Conversations: Multimodal Context for Natural Question and Response Generation

Nasrin Mostafazadeh, Chris Brockett, Bill Dolan, Michel Galley, Jianfeng Gao, Georgios P. Spithourakis, Lucy Vanderwende

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.03218 2017-03-16 cs.AI cs.CL cs.LG cs.MA 62%

Learning to Play Guess Who? and Inventing a Grounded Language as a Consequence

Emilio Jorge, Mikael Kågebäck, Fredrik D. Johansson, Emil Gustavsson

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments Previous version was accepted to Deep Reinforcement Learning Workshop at NIPS 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.07182 2017-03-07 cs.CL cs.CV cs.GT cs.LG cs.MA 62%

Multi-Agent Cooperation and the Emergence of (Natural) Language

Angeliki Lazaridou, Alexander Peysakhovich, Marco Baroni

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Comments Accepted at ICLR 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1512.00355 2015-12-02 cs.AI cs.LG 62%

Taxonomy grounded aggregation of classifiers with different label sets

Amrita Saha, Sathish Indurthi, Shantanu Godbole, Subendhu Rongali, Vikas C. Raykar

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments Under review by AISTATS 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1508.06161 2015-08-26 cs.RO cs.AI cs.CL cs.HC cs.LG 62%

Robot Language Learning, Generation, and Comprehension

Daniel Paul Barrett, Scott Alan Bronikowski, Haonan Yu, Jeffrey Mark Siskind

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1401.5390 2014-01-22 cs.CL cs.AI cs.LG 62%

Learning to Win by Reading Manuals in a Monte-Carlo Framework

S. R. K. Branavan, David Silver, Regina Barzilay

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Journal ref Journal Of Artificial Intelligence Research, Volume 43, pages 661-704, 2012

详情

展开后加载摘要…

URL PDF HTML 收藏
1005.5253 2010-05-31 cs.CL cs.AI cs.HC cs.LG 62%

Using Soft Constraints To Learn Semantic Models Of Descriptions Of Shapes

Sergio Guadarrama, David P. Pancho

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments 8 pages, 8 figures, WCCI'10 Conference

详情

展开后加载摘要…

URL PDF HTML 收藏