机构
*
MoE Key Lab of Artificial Intelligence, Institute of AI, Shanghai Jiao Tong University(教育部人工智能重点实验室,上海交通大学人工智能研究院)
;
Central Research Institute, Huawei(华为中央研究院)
CAC-VLA: Context-Gated Action Conditioning for Vision-Language-Action Models
CAC-VLA:用于视觉-语言-动作模型的上下文门控动作条件调节
Yifu Xiong, Wenhao Yu, Jiaxuan Lin, Bojun Zou, Jiahao Li, Lu Zhang, Yanyong Zhang, Jianmin Ji
机构
*
University of Science and Technology of China (USTC)(中国科学技术大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning
通过自主经验探索与事后经验利用赋能GUI智能体任务规划
Tianyi Men, Zhuoran Jin, Pengfei Cao, Yubo Chen, Kang Liu, Jun Zhao
机构
*
The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所复杂系统认知与决策智能重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
机构
*
Nanyang Technological University(南洋理工大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
A*STAR Institute for Infocomm Research (I2R)(新加坡科技研究局资讯通信研究院)
Action QFormer: Structured Representation Shaping under Action Supervision in Vision-Language-Action Models
动作QFormer:视觉-语言-动作模型中动作监督下的结构化表示塑造
Yufeng Ji, Wenhao Tang, Haoyi Niu, Koushil Sreenath, Yi Wu, Zhongyu Li
机构
*
Shanghai Qizhi Institute(上海期智研究院)
;
The Chinese University of Hong Kong(香港中文大学)
;
Hong Kong Embodied AI Lab(香港具身人工智能实验室)
;
Tsinghua University(清华大学)
;
University of California, Berkeley(加州大学伯克利分校)
Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review
用于无人机机器人和双手操作的视觉语言动作(VLA)模型综述
Inkyu Sa, Chanoh Park, Hea-Min Lee, Donghee Noh, Ho Seok Ahn
机构
*
Chef Robotics
;
RovifyLab
;
IT Application Research Center, Jeonbuk Regional Branch
;
Department of Electrical, Computer and Software Engineering(电气与计算机软件工程系)