Co-VLA: Coordination-Aware Structured Action Modeling for Dual-Arm Vision-Language-Action Systems
Co-VLA:面向双臂视觉-语言-动作系统的协调感知结构化动作建模
机构 * Donghua University(东华大学) ; Samsung R&D Institute China-Beijing (SRCB)(三星中国北京研究院) ; Samsung AI Center, DS Division(三星DS部门AI中心)
专题命中 VLA模型 :VLA(title,title_cn);vision-language-action(title,abstract);action model(title);分类 cs.RO
AI总结 针对双臂紧耦合任务中隐式协调不足的问题,提出Co-VLA框架,通过结构化动作专家和潜在感知控制器显式引入协调先验,在仿真和真实场景中显著提升成功率和效率。