arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

VLA / 视觉-语言-动作模型

视觉-语言-动作模型、机器人基础模型和语言条件机器人控制。

共收录 9819 信号源:cs.RO, cs.CV, cs.AI, cs.LG

1. VLA模型 9146 篇

2308.07751 2023-08-16 cs.CV 57%

CASPNet++: Joint Multi-Agent Motion Prediction

Maximilian Schäfer, Kun Zhao, Anton Kummert

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.05382 2023-08-11 cs.CV 57%

Interaction-aware Joint Attention Estimation Using People Attributes

Chihiro Nakatani, Hiroaki Kawashima, Norimichi Ukita

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Accepted to ICCV2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.05046 2023-08-10 cs.CL cs.LG 57%

RadGraph2: Modeling Disease Progression in Radiology Reports via Hierarchical Information Extraction

Sameer Khanna, Adam Dejl, Kibo Yoon, Quoc Hung Truong, Hanh Duong, Agustina Saenz, Pranav Rajpurkar

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Comments Accepted at Machine Learning for Healthcare 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.01001 2023-08-03 cs.RO 57%

Push to know! -- Visuo-Tactile based Active Object Parameter Inference with Dual Differentiable Filtering

Anirvan Dutta, Etienne Burdet, Mohsen Kaboli

专题命中 VLA模型 :action model(abstract);分类 cs.RO

Comments 8 pages. Accepted at IROS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.12490 2023-08-03 cs.CL cs.AI 57%

Improve Event Extraction via Self-Training with Gradient Guidance

Zhiyang Xu, Jay-Yoon Lee, Lifu Huang

专题命中 VLA模型 :action model(abstract);分类 cs.AI

Comments ACL 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.03971 2023-08-02 cs.CV 57%

Universal Prototype Transport for Zero-Shot Action Recognition and Localization

Pascal Mettes

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Journal ref International Journal of Computer Vision (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.00153 2023-08-01 cs.RO cs.SY eess.SY 57%

Probabilistic Traversability Model for Risk-Aware Motion Planning in Off-Road Environments

Xiaoyi Cai, Michael Everett, Lakshay Sharma, Philip R. Osteen, Jonathan P. How

专题命中 VLA模型 :action model(abstract);分类 cs.RO

Comments To appear in IROS23. Video and code: https://github.com/mit-acl/mppi_numba

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.14368 2023-07-28 cs.FL cs.AI 57%

Synthesis of Procedural Models for Deterministic Transition Systems

Javier Segovia-Aguas, Jonathan Ferrer-Mestres, Sergio Jiménez

专题命中 VLA模型 :action model(abstract);分类 cs.AI

Comments Conference paper accepted at ECAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.02677 2023-07-07 cs.CV 57%

Caption Anything: Interactive Image Description with Diverse Multimodal Controls

Teng Wang, Jinrui Zhang, Junjie Fei, Hao Zheng, Yunlong Tang, Zhe Li, Mingqi Gao, Shanshan Zhao

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Tech-report

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.02090 2023-07-06 cs.CV 57%

Interactive Conversational Head Generation

Mohan Zhou, Yalong Bai, Wei Zhang, Ting Yao, Tiejun Zhao

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments arXiv admin note: text overlap with arXiv:2112.13548

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.06258 2023-07-04 cs.AI cs.CL 57%

A model of interaction semantics

Johannes Reich

专题命中 VLA模型 :action model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.12050 2023-06-22 cs.CV 57%

Analyzing Font Style Usage and Contextual Factors in Real Images

Naoya Yasukochi, Hideaki Hayashi, Daichi Haraguchi, Seiichi Uchida

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Accepted at ICDAR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.10042 2023-06-21 cs.IR cs.AI cs.CL 57%

A Pairing Enhancement Approach for Aspect Sentiment Triplet Extraction

Fan Yang, Mian Zhang, Gongzhen Hu, Xiabing Zhou

专题命中 VLA模型 :action model(abstract);分类 cs.AI

Comments 12 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.04796 2023-06-13 cond-mat.mtrl-sci cs.LG physics.plasm-ph 57%

Physics-separating artificial neural networks for predicting initial stages of Al sputtering and thin film deposition in Ar plasma discharges

Tobias Gergs, Thomas Mussenbrock, Jan Trieschmann

专题命中 VLA模型 :action model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.04887 2023-06-09 cs.AI cs.NI 57%

Big-data-driven and AI-based framework to enable personalization in wireless networks

Rawan Alkurd, Ibrahim Abualhaol, Halim Yanikomeroglu

专题命中 VLA模型 :action model(abstract);分类 cs.AI

Journal ref IEEE Communications Magazine ( Volume: 58, Issue: 3, March 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.01475 2023-06-05 cs.IR cs.LG 57%

Prompt Tuning Large Language Models on Personalized Aspect Extraction for Recommendations

Pan Li, Yuyan Wang, Ed H. Chi, Minmin Chen

专题命中 VLA模型 :action model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.15006 2023-06-01 cs.CY cs.AI 57%

A Human-in-the-Loop Approach for Information Extraction from Privacy Policies under Data Scarcity

Michael Gebauer, Faraz Maschhur, Nicola Leschke, Elias Grünewald, Frank Pallas

专题命中 VLA模型 :action model(abstract);分类 cs.AI

Comments Accepted for 2023 IEEE European Symposium on Security and Privacy Workshops (EuroS&P)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.15603 2023-05-26 cs.LG physics.flu-dyn 57%

Learning Lagrangian Fluid Mechanics with E($3$)-Equivariant Graph Neural Networks

Artur P. Toshev, Gianluca Galletti, Johannes Brandstetter, Stefan Adami, Nikolaus A. Adams

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Comments GSI'23 6th International Conference on Geometric Science of Information; 10 pages; oral. arXiv admin note: substantial text overlap with arXiv:2304.00150

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.14612 2023-05-25 cs.CV stat.AP 57%

Assessment of Anterior Cruciate Ligament Injury Risk Based on Human Key Points Detection Algorithm

Ziyu Gong, Xiong Zhao, Chen Yang

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments 17 pages,and 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.13262 2023-05-23 cs.SD cs.LG eess.AS 57%

Modulation Extraction for LFO-driven Audio Effects

Christopher Mitcheltree, Christian J. Steinmetz, Marco Comunità, Joshua D. Reiss

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Comments Accepted to DAFx 2023. Listening samples and plugins can be found at https://christhetree.github.io/mod_extraction/

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.09212 2023-05-17 eess.AS cs.CV cs.MM cs.SD 57%

Cross-Modal Global Interaction and Local Alignment for Audio-Visual Speech Recognition

Yuchen Hu, Ruizhe Li, Chen Chen, Heqing Zou, Qiushi Zhu, Eng Siong Chng

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments 12 pages, 5 figures, Accepted by IJCAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.07138 2023-04-14 cs.AI cs.CL 57%

Integrating AI Planning with Natural Language Processing: A Combination of Explicit and Tacit Knowledge

Kebing Jin, Hankz Hankui Zhuo

专题命中 VLA模型 :action model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.02714 2023-04-07 cs.CV eess.SP 57%

Learning Stage-wise GANs for Whistle Extraction in Time-Frequency Spectrograms

Pu Li, Marie Roch, Holger Klinck, Erica Fleishman, Douglas Gillespie, Eva-Marie Nosal, Yu Shiu, Xiaobai Liu

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions of Multimedia (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.10434 2023-04-06 cs.CV cs.MM 57%

Frame-wise Cross-modal Matching for Video Moment Retrieval

Haoyu Tang, Jihua Zhu, Meng Liu, Zan Gao, Zhiyong Cheng

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments 12 pages; accepted by IEEE TMM

Journal ref IEEE Transactions on Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.00150 2023-04-04 cs.LG physics.flu-dyn 57%

E($3$) Equivariant Graph Neural Networks for Particle-Based Fluid Mechanics

Artur P. Toshev, Gianluca Galletti, Johannes Brandstetter, Stefan Adami, Nikolaus A. Adams

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Comments ICLR 2023 Workshop on Physics for Machine Learning

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.10876 2023-03-28 cs.CV cs.MA 57%

EqMotion: Equivariant Multi-agent Motion Prediction with Invariant Interaction Reasoning

Chenxin Xu, Robby T. Tan, Yuhong Tan, Siheng Chen, Yu Guang Wang, Xinchao Wang, Yanfeng Wang

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Accepted to CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.10904 2023-03-22 cs.CV 57%

Actionlet-Dependent Contrastive Learning for Unsupervised Skeleton-Based Action Recognition

Lilang Lin, Jiahang Zhang, Jiaying Liu

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Accepted by CVPR2023 (Highlight). The project page is at https://langlandslin.github.io/projects/ActCLR/

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.02078 2023-03-21 cs.CL cs.AI 57%

FGSI: Distant Supervision for Relation Extraction method based on Fine-Grained Semantic Information

Chenghong Sun, Weidong Ji, Guohui Zhou, Hui Guo, Zengxiang Yin, Yuqi Yue

专题命中 VLA模型 :action model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.07202 2023-03-14 cs.AI math.OC 57%

Optimization of the location and design of urban green spaces

Caroline Leboeuf, Margarida Carvalho, Yan Kestens, Benoît Thierry

专题命中 VLA模型 :action model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.02073 2023-03-06 cs.LG 57%

How To Guide Your Learner: Imitation Learning with Active Adaptive Expert Involvement

Xu-Hui Liu, Feng Xu, Xinyu Zhang, Tianyuan Liu, Shengyi Jiang, Ruifeng Chen, Zongzhang Zhang, Yang Yu

专题命中 VLA模型 :action model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏