arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

视觉大模型 / VLM

视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。

共收录 7387 信号源:cs.CV, cs.AI, cs.LG

1. 视觉定位与Grounding 7387 篇

2102.05126 2022-01-17 cs.CL cs.AI cs.LG 62%

AuGPT: Auxiliary Tasks and Data Augmentation for End-To-End Dialogue with Pre-Trained Language Models

Jonáš Kulhánek, Vojtěch Hudeček, Tomáš Nekvinda, Ondřej Dušek

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Journal ref Proceedings of the 3rd Workshop on Natural Language Processing for Conversational AI (2021), 198-210

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.11735 2022-01-11 cs.LG cs.AI cs.LO 62%

Lifted Model Checking for Relational MDPs

Wen-Chi Yang, Jean-François Raskin, Luc De Raedt

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.14115 2022-01-02 cs.CV cs.AI 62%

Visually Grounded Concept Composition

Bowen Zhang, Hexiang Hu, Linlu Qiu, Peter Shaw, Fei Sha

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments Findings of EMNLP 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.00534 2021-12-30 cs.CV cs.AI cs.CL cs.RO 62%

TEACh: Task-driven Embodied Agents that Chat

Aishwarya Padmakumar, Jesse Thomason, Ayush Shrivastava, Patrick Lange, Anjali Narayan-Chen, Spandana Gella, Robinson Piramuthu, Gokhan Tur, Dilek Hakkani-Tur

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments Accepted at AAAI 2022; 7 pages main, 28 pages total, 29 figures; Version 3 uses a new test set for EDH instances that restrict evaluation to state changes only on task-relevant objects

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.13758 2021-12-28 cs.CL cs.AI cs.LG cs.RO 62%

Bridging the Gap: Using Deep Acoustic Representations to Learn Grounded Language from Percepts and Raw Speech

Gaoussou Youssouf Kebe, Luke E. Richards, Edward Raff, Francis Ferraro, Cynthia Matuszek

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments To appear in the Proceedings of the 36th AAAI Conference on Artificial Intelligence. February 2022, Vancouver

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.08268 2021-12-16 cs.LG cs.AI cs.CY 62%

Prescriptive Machine Learning for Automated Decision Making: Challenges and Opportunities

Eyke Hüllermeier

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.00311 2021-12-01 stat.ML cs.AI cs.LG 62%

What's a good imputation to predict with missing values?

Marine Le Morvan, Julie Josse, Erwan Scornet, Gaël Varoquaux

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.01115 2021-11-02 cs.RO cs.AI cs.LG 62%

Learning Language-Conditioned Robot Behavior from Offline Data and Crowd-Sourced Annotation

Suraj Nair, Eric Mitchell, Kevin Chen, Brian Ichter, Silvio Savarese, Chelsea Finn

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments Conference on Robot Learning (CoRL) 2021. 24 Pages, 18 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.09150 2021-10-22 cs.CL cs.AI cs.CV 62%

VisualSem: A High-quality Knowledge Graph for Vision and Language

Houda Alberts, Teresa Huang, Yash Deshpande, Yibo Liu, Kyunghyun Cho, Clara Vania, Iacer Calixto

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments Accepted for publication at the 1st Multilingual Representation Learning workshop (MRL 2021) co-located with EMNLP 2021. 15 pages, 8 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.10649 2021-09-23 cs.CV cs.AI 62%

Caption Enriched Samples for Improving Hateful Memes Detection

Efrat Blaier, Itzik Malkiel, Lior Wolf

专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI

Comments EMNLP 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.06304 2021-08-23 cs.AI cs.CL cs.CV 62%

What is Multimodality?

Letitia Parcalabescu, Nils Trost, Anette Frank

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments Paper accepted for publication at MMSR 2021; 10 pages, 5 figures

Journal ref Proceedings of the 1st Workshop on Multimodal Semantic Representations (MMSR), 2021, Groningen, Netherlands (Online), Association for Computational Linguistics, p. 1--10

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.07253 2021-08-18 cs.CV cs.CL cs.LG 62%

Who's Waldo? Linking People Across Text and Images

Claire Yuqing Cui, Apoorv Khandelwal, Yoav Artzi, Noah Snavely, Hadar Averbuch-Elor

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Comments Published in ICCV 2021 (Oral). Project webpage: https://whoswaldo.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.14593 2021-08-02 cs.CL cs.AI cs.LG cs.RO 62%

Neural Variational Learning for Grounded Language Acquisition

Nisha Pillai, Cynthia Matuszek, Francis Ferraro

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Journal ref 2021 30th IEEE International Conference on Robot and Human Interactive Communication (RO-MAN)

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.04887 2021-07-15 cs.LG cs.AI stat.ML 62%

Interaction-Grounded Learning

Tengyang Xie, John Langford, Paul Mineiro, Ida Momennejad

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments Published in ICML 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.04598 2021-06-22 cs.SD cs.CV cs.LG eess.AS eess.IV 62%

Cross-Modal learning for Audio-Visual Video Parsing

Jatin Lamba, Abhishek, Jayaprakash Akula, Rishabh Dabral, Preethi Jyothi, Ganesh Ramakrishnan

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Comments Work accepted at Interspeech 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.08143 2021-06-22 q-bio.NC cs.AI cs.LG cs.NE 62%

Constrained plasticity reserve as a natural way to control frequency and weights in spiking neural networks

Oleg Nikitin, Olga Lukyanova, Alex Kunin

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments 24 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.04244 2021-06-09 cs.AI cs.LG 62%

Counterfactuals and Causability in Explainable Artificial Intelligence: Theory, Algorithms, and Applications

Yu-Liang Chou, Catarina Moreira, Peter Bruza, Chun Ouyang, Joaquim Jorge

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.00920 2021-06-03 cs.CL cs.AI cs.LG 62%

DialoGraph: Incorporating Interpretable Strategy-Graph Networks into Negotiation Dialogues

Rishabh Joshi, Vidhisha Balachandran, Shikhar Vashishth, Alan Black, Yulia Tsvetkov

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments Accepted at ICLR 2021; https://openreview.net/forum?id=kDnal_bbb-E

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.01981 2021-06-02 cs.CL cs.AI cs.CV 62%

Open Domain Dialogue Generation with Latent Images

Ze Yang, Wei Wu, Huang Hu, Can Xu, Wei Wang, Zhoujun Li

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments AAAI2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.07168 2021-05-20 cs.LG cs.AI econ.EM stat.ML 62%

Cohort Shapley value for algorithmic fairness

Masayoshi Mase, Art B. Owen, Benjamin B. Seiler

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.02936 2021-05-10 cs.LG cs.AI cs.MS stat.ML 62%

Exact Acceleration of K-Means++ and K-Means$\|$

Edward Raff

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments to appear in the 30th International Joint Conference on Artificial Intelligence (IJCAI-21)

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.06278 2021-04-23 cs.CV cs.AI 62%

COSMOS: Catching Out-of-Context Misinformation with Self-Supervised Learning

Shivangi Aneja, Chris Bregler, Matthias Nießner

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments Video : https://youtu.be/riI3Cl2xy10

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.01766 2021-04-20 cs.CL cs.AI cs.LG 62%

Word meaning in minds and machines

Brenden M. Lake, Gregory L. Murphy

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments In press at Psychological Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.03775 2021-04-01 cs.CV cs.AI 62%

Text-to-Image Generation Grounded by Fine-Grained User Attention

Jing Yu Koh, Jason Baldridge, Honglak Lee, Yinfei Yang

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments To appear in WACV 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.03431 2021-01-12 cs.AI cs.CL cs.CV cs.RO 62%

Are We There Yet? Learning to Localize in Embodied Instruction Following

Shane Storks, Qiaozi Gao, Govind Thattai, Gokhan Tur

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI

Comments Accepted to HAI @ AAAI 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.05613 2020-12-10 cs.LG cs.AI 62%

Planning with Learned Object Importance in Large Problem Instances using Graph Neural Networks

Tom Silver, Rohan Chitnis, Aidan Curtis, Joshua Tenenbaum, Tomas Lozano-Perez, Leslie Pack Kaelbling

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments AAAI 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.09529 2020-12-10 cs.LG cs.CV stat.ML 62%

Learning Activation Functions: A new paradigm for understanding Neural Networks

Mohit Goyal, Rajan Goyal, Brejesh Lall

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Comments A modified version of the article has been published in IEEE WCCI 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.02659 2020-12-07 cs.AI cs.LG cs.NE 62%

Understanding Attention: In Minds and Machines

Shriraj P. Sawant, Shruti Singh

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

Comments Accepted at NeurIPS 2020 Workshop: ML Retrospectives, Surveys & Meta-Analyses (ML-RSA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.14204 2020-12-01 cs.CV cs.LG stat.ML 62%

Class-agnostic Object Detection

Ayush Jaiswal, Yue Wu, Pradeep Natarajan, Premkumar Natarajan

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG

Comments To appear in Proceedings of WACV 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.09854 2020-11-20 cs.LG cs.AI 62%

Generalized Inverse Planning: Learning Lifted non-Markovian Utility for Generalizable Task Representation

Sirui Xie, Feng Gao, Song-Chun Zhu

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏