arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

VLA / 视觉-语言-动作模型

视觉-语言-动作模型、机器人基础模型和语言条件机器人控制。

共收录 9819 信号源:cs.RO, cs.CV, cs.AI, cs.LG

1. VLA模型 9146 篇

2112.09836 2023-07-10 cs.AI cs.LG 62%

Creativity of AI: Hierarchical Planning Model Learning for Facilitating Deep Reinforcement Learning

Hankz Hankui Zhuo, Shuting Deng, Mu Jin, Zhihao Ma, Kebing Jin, Chen Chen, Chao Yu

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.05615 2023-05-30 cs.CV cs.AI 62%

FIGO: Enhanced Fingerprint Identification Approach Using GAN and One Shot Learning Techniques

Ibrahim Yilmaz, Mahmoud Abouyoussef

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.15589 2023-05-15 cs.LG cs.AI 62%

Inapplicable Actions Learning for Knowledge Transfer in Reinforcement Learning

Leo Ardon, Alberto Pozanco, Daniel Borrajo, Sumitra Ganesh

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.02412 2023-05-09 cs.CL cs.AI cs.LG 62%

Plan, Eliminate, and Track -- Language Models are Good Teachers for Embodied Agents

Yue Wu, So Yeon Min, Yonatan Bisk, Ruslan Salakhutdinov, Amos Azaria, Yuanzhi Li, Tom Mitchell, Shrimai Prabhumoye

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.12484 2023-05-03 cs.CV cs.AI 62%

DocParser: End-to-end OCR-free Information Extraction from Visually Rich Documents

Mohamed Dhouib, Ghassen Bettaieb, Aymen Shabou

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.AI

Comments The 17th International Conference on Document Analysis and Recognition

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.01765 2023-03-22 math.NA cs.AI cs.LG cs.NA physics.comp-ph 62%

opPINN: Physics-Informed Neural Network with operator learning to approximate solutions to the Fokker-Planck-Landau equation

Jae Yong Lee, Juhi Jang, Hyung Ju Hwang

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

Comments 28 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.01478 2023-02-13 cs.AI cs.LG 62%

Clustered Embedding Learning for Recommender Systems

Yizhou Chen, Guangda Huzhang, Anxiang Zeng, Qingtao Yu, Hui Sun, Heng-yi Li, Jingyi Li, Yabo Ni, Han Yu, Zhiming Zhou

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.11096 2023-01-31 cs.LG cs.AI cs.NE 62%

Let Offline RL Flow: Training Conservative Agents in the Latent Space of Normalizing Flows

Dmitriy Akimov, Vladislav Kurenkov, Alexander Nikulin, Denis Tarasov, Sergey Kolesnikov

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

Comments Accepted at 3rd Offline Reinforcement Learning Workshop at Neural Information Processing Systems, 2022. Source code: https://github.com/tinkoff-ai/cnf

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.03091 2023-01-26 cs.CL cs.AI cs.DB cs.IR cs.LG 62%

Relation Adversarial Network for Low Resource Knowledge Graph Completion

Ningyu Zhang, Shumin Deng, Zhanlin Sun, Jiaoayan Chen, Wei Zhang, Huajun Chen

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

Comments WWW2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.04704 2023-01-20 cs.CV cs.AI 62%

An Efficient Pattern Mining Convolution Neural Network (CNN) algorithm with Grey Wolf Optimization (GWO)

Aatif Jamshed, Bhawna Mallick, Rajendra Kumar Bharti

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.AI

Journal ref The Imaging Science Journal 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.14823 2022-10-31 cs.CV cs.AI 62%

Visual Answer Localization with Cross-modal Mutual Knowledge Transfer

Yixuan Weng, Bin Li

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.AI

Comments 4 pages, 3 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.10983 2022-10-14 cs.IR cs.AI cs.LG 62%

Online POI Recommendation: Learning Dynamic Geo-Human Interactions in Streams

Dongjie Wang, Kunpeng Liu, Hui Xiong, Yanjie Fu

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.01799 2022-10-06 cs.LG cs.AI 62%

STGIN: A Spatial Temporal Graph-Informer Network for Long Sequence Traffic Speed Forecasting

Ruikang Luo, Yaofeng Song, Liping Huang, Yicheng Zhang, Rong Su

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

Comments 12 pages, 18 figures and 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.12561 2022-09-27 cs.IR cs.CV cs.LG 62%

Improving Document Image Understanding with Reinforcement Finetuning

Bao-Sinh Nguyen, Dung Tien Le, Hieu M. Vu, Tuan Anh D. Nguyen, Minh-Tien Nguyen, Hung Le

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.LG

Comments Accepted to ICONIP 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.03648 2022-08-09 cs.CV cs.AI eess.IV 62%

Weakly Supervised Online Action Detection for Infant General Movements

Tongyi Luo, Jia Xiao, Chuncao Zhang, Siheng Chen, Yuan Tian, Guangjun Yu, Kang Dang, Xiaowei Ding

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.AI

Comments MICCAI 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.09289 2022-06-24 cs.LG cs.AI 62%

MHNF: Multi-hop Heterogeneous Neighborhood information Fusion graph representation learning

Yundong Sun, Dongjie Zhu, Haiwen Du, Zhaoshuo Tian

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

Comments update some content of the paper and polish the language!

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.00110 2022-06-17 cs.AI cs.LG 62%

Classical Planning in Deep Latent Space

Masataro Asai, Hiroshi Kajino, Alex Fukunaga, Christian Muise

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

Comments Accepted in Journal of Artificial Intelligence Research (JAIR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01088 2022-06-06 eess.IV cs.CV cs.LG 62%

Machine Learning-based Lung and Colon Cancer Detection using Deep Feature Extraction and Ensemble Learning

Md. Alamin Talukder, Md. Manowarul Islam, Md Ashraf Uddin, Arnisha Akhter, Khondokar Fida Hasan, Mohammad Ali Moni

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.LG

Comments Accepted for publication in the Special Issue of Expert Systems with Applications (IF:6.954, Cite:12.70) How to Cite: Md. Alamin Talukder, Md. Manowarul Islam, Md Ashraf Uddin, Arnisha Akhter, Khondokar Fida Hasan, Mohammad Ali Moni. "Machine Learning-based Lung and Colon Cancer Detection using Deep Feature Extraction and Ensemble Learning", Expert Systems with Applications. 2022 Jun 1

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.15025 2022-05-31 cs.CV cs.AI cs.CL 62%

An Efficient Modern Baseline for FloodNet VQA

Aditya Kane, Sahil Khose

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.AI

Comments Under review, 4 pages, 2 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.02490 2022-05-24 cs.CL cs.AI cs.LG 62%

FastRE: Towards Fast Relation Extraction with Convolutional Encoder and Improved Cascade Binary Tagging Framework

Guozheng Li, Xu Chen, Peng Wang, Jiafeng Xie, Qiqing Luo

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

Comments Accepted to IJCAI-ECAI 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.05368 2022-05-12 cs.CL cs.AI cs.HC cs.LG 62%

Pre-trained Language Models as Re-Annotators

Chang Shu

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

Comments Thesis of Master of Science by Research (M.Res) in Linguistics with Distinction; University of Edinburgh, 2022; 107 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.10621 2022-03-22 cs.CL cs.AI cs.LG 62%

Immersive Text Game and Personality Classification

Wanshui Li, Yifan Bai, Jiaxuan Lu, Kexin Yi

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.10090 2022-03-22 cs.CV cs.LG 62%

FaceMap: Towards Unsupervised Face Clustering via Map Equation

Xiaotian Yu, Yifan Yang, Aibo Wang, Ling Xing, Hanling Yi, Guangming Lu, Xiaoyu Wang

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.01883 2022-03-04 eess.IV cs.CV cs.LG 62%

ROCT-Net: A new ensemble deep convolutional model with improved spatial resolution learning for detecting common diseases from retinal OCT images

Mohammad Rahimzadeh, Mahmoud Reza Mohammadi

专题命中 VLA模型 :action model(abstract);分类 cs.CV、cs.LG

Comments This is a preprint of an article published in the ICCKE 2021 conference. The final authenticated version is available online at https://doi.org/10.1109/ICCKE54056.2021.9721471. The code of this paper is shared at https://github.com/mr7495/OCT-classification

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.08373 2022-02-21 cs.LG cs.AI cs.CL 62%

Text-Based Action-Model Acquisition for Planning

Kebing Jin, Huaixun Chen, Hankz Hankui Zhuo

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.03199 2022-01-11 cs.RO cs.AI 62%

Task planning and explanation with virtual actions

Guowei Cui, Xiaoping Chen

专题命中 VLA模型 :action model(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.15147 2021-10-01 cs.AI cs.LG 62%

Reinforcement Learning with Information-Theoretic Actuation

Elliot Catt, Marcus Hutter, Joel Veness

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.06069 2021-08-16 cs.IR cs.AI cs.CL cs.LG 62%

Zero-shot Task Transfer for Invoice Extraction via Class-aware QA Ensemble

Prithiviraj Damodaran, Prabhkaran Singh, Josemon Achankuju

专题命中 VLA模型 :action model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.03298 2021-08-10 cs.CV cs.RO 62%

Bifold and Semantic Reasoning for Pedestrian Behavior Prediction

Amir Rasouli, Mohsen Rohani, Jun Luo

专题命中 VLA模型 :action model(abstract);分类 cs.RO、cs.CV

Comments ICCV 2021. 11 pages; 5 Figures; 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.04538 2021-07-12 cs.RO cs.AI 62%

Learning Interaction-aware Guidance Policies for Motion Planning in Dense Traffic Scenarios

Bruno Brito, Achin Agarwal, Javier Alonso-Mora

专题命中 VLA模型 :action model(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏