FD-VLA: Force-Distilled Vision-Language-Action Model for Contact-Rich Manipulation
FD-VLA:力感知的视觉-语言-动作模型用于接触密集的 manipulation
Ruiteng Zhao, Wenshuo Wang, Yicheng Ma, Xiaocong Li, Francis E. H. Tay, Marcelo H. Ang, Haiyue Zhu
机构
*
Advanced Robotics Centre, National University of Singapore(新加坡国立大学先进机器人中心)
;
School of Electrical & Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院)
;
College of Information Science and Technology, Eastern Institute of Technology(东部技术学院信息科学与技术学院)
;
John A. Paulson School of Engineering and Applied Sciences, Harvard University(哈佛大学约翰·A·保罗森工程与应用科学学院)
;
Advanced Robotics Centre at National University of Singapore(新加坡国立大学先进机器人中心)
;
Singapore Institute of Manufacturing Technology, Agency for Science, Technology and Research (A*STAR)(新加坡制造技术研究所,科技研究局(A*STAR))