MPCoT: Reward-Guided Multi-Path Latent Reasoning for Test-Time Scalable Vision-Language-Action
MPCoT: 奖励引导的多路径潜在推理用于测试时可扩展的视觉-语言-动作
Boyang Zhang, Lianlei Shan
机构
*
Department of Electrical and Computer Engineering, Boston University(波士顿大学电气与计算机工程系)
;
Department of Computer Science, Tsinghua University(清华大学计算机系)
OrthoSkillVLA: Continual Skill Learning via Gradient-Informed Skill Subspace Adaptation
OrthoSkillVLA:通过梯度感知技能子空间自适应实现持续技能学习
Jiaqi Wang, Zhou Fang, Qiongfeng Shi, Yi Zhou
机构
*
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
School of Electronic Science and Engineering, Southeast University(东南大学电子科学与工程学院)
Comments8 pages, 5 figures. Introduces DECOWAM, a decoupled whole-body world-action model for legged mobile manipulation, and the ARMDOG real-robot dataset
Towards Surgical World-Action Modeling: A Preliminary Joint Visual-Trajectory Forecasting for Surgical Motion Planning
面向外科世界动作建模:外科手术运动规划的初步联合视觉-轨迹预测
Weiliang Huang, Huanrong Liu, Bob Zhang, Qi Dou, Zhen Chen, Yun Gu, Guy Rosman, Qingbiao Li
机构
*
University of Macau(澳门大学)
;
University of Macau Advanced Research Institute in Hengqin(澳门大学横琴研究院)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Tongren Hospital(同仁医院)
;
Duke University(杜克大学)
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
ImprintX Robotics(ImprintX机器人公司)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)