机构
*
School of Mechanical and Electronic Engineering, Wuhan University of Technology(武汉理工大学机电工程学院)
;
Intelligent Transportation Systems Research Center, Wuhan University of Technology(武汉理工大学智能交通系统研究中心)
;
School of Computer Science and Artificial Intelligence, Wuhan University of Technology(武汉理工大学计算机科学与人工智能学院)
vla.cpp: A Unified Inference Runtime for Vision-Language-Action Models
vla.cpp:视觉-语言-动作模型的统一推理运行时
Khanh D. Nguyen, Hung T. Ho, Chinh T. Nguyen, Thanh Q. Duong, Linh D. Le, Duy M. H. Nguyen, Vien A. Ngo, An T. Le
机构
*
VinRobotics
;
Center for AI Research, VinUniversity(VinUniversity 人工智能研究中心)
;
Intelligent Autonomous Systems, TU Darmstadt(达姆施塔特工业大学智能自主系统)
;
Max Planck Research School for Intelligent Systems(马克斯·普朗克智能系统研究学院)
;
University of Stuttgart(斯图加特大学)
;
German Research Center for Artificial Intelligence(德国人工智能研究中心)
机构
*
School of Computer Science and Technology, East China Normal University(华东师范大学计算机科学与技术学院)
;
State Key Laboratory of Submarine Geoscience, School of Automation and Intelligent Sensing, Shanghai Jiao Tong University(上海交通大学自动化与智能感知学院海底科学国家重点实验室)
G$^3$VLA: Geometric inductive bias for Vision-Language-Action Models
G$^3$VLA:视觉-语言-动作模型的几何归纳偏置
Yue Peng, Yongzhe Zhao, Artur Habuda, Khuyen Pham, Yanheng Zhu, Tran Nguyen Le, Fares Abu-Dakka, Li Guo
机构
*
New York University Shanghai(上海纽约大学)
;
Technical University of Denmark(丹麦技术大学)
;
MBZUAI - Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
New York University Abu Dhabi(纽约大学阿布扎比分校)
CosFly-VLA: A Spatially Aware Vision-Language-Action Model for UAV Tracking
CosFly-VLA:一种用于无人机跟踪的空间感知视觉-语言-动作模型
Ruilong Ren, Songsheng Cheng, Yunpeng Zhou, Hanxuan Chen, Xiangyue Wang, Tianle Zeng, Shuai Yuan, Binbo Li, Hanzhong Guo, Ji Pei, Da Zhang, Kangli Wang
机构
*
Autel Robotics(大疆创新科技有限公司)
;
Northeast Normal University(东北师范大学)
;
Southern University of Science and Technology(南方科技大学)
;
Peking University(北京大学)
;
University of Hong Kong(香港大学)
CAC-VLA: Context-Gated Action Conditioning for Vision-Language-Action Models
CAC-VLA:用于视觉-语言-动作模型的上下文门控动作条件调节
Yifu Xiong, Wenhao Yu, Jiaxuan Lin, Bojun Zou, Jiahao Li, Lu Zhang, Yanyong Zhang, Jianmin Ji
机构
*
University of Science and Technology of China (USTC)(中国科学技术大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
Revisiting Parameter Redundancy in Vision-Language-Action Models: Insights from VLM-to-VLA Adaptation
重新审视视觉-语言-动作模型中的参数冗余:从VLM到VLA适配的见解
Fengnian Zhang, Tao Huang, Siyu Xu, Zhong Jin, Chang Xu
机构
*
Computer Network Information Center, Chinese Academy of Sciences(中国科学院计算机网络信息中心)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学与工程学院)
;
School of Computer Science, The University of Sydney(悉尼大学计算机科学学院)
机构
*
Brain-inspired Cognitive AI Lab, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所类脑认知智能实验室)
;
Beijing Key Laboratory of Safe AI and Superalignment(北京市安全人工智能与超级对齐重点实验室)
;
Beijing Institute of AI Safety and Governance(北京人工智能安全与治理研究所)
;
Gaoling School of AI, Renmin University of China(中国人民大学高瓴人工智能学院)
;
University of Chinese Academy of Sciences (UCAS)(中国科学院大学)
Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review
用于无人机机器人和双手操作的视觉语言动作(VLA)模型综述
Inkyu Sa, Chanoh Park, Hea-Min Lee, Donghee Noh, Ho Seok Ahn
机构
*
Chef Robotics
;
RovifyLab
;
IT Application Research Center, Jeonbuk Regional Branch
;
Department of Electrical, Computer and Software Engineering(电气与计算机软件工程系)
专题命中
VLA模型
:VLA(title,title_cn);vision language action(title,abstract);分类 cs.RO、cs.AI、cs.LG