FD-VLA: Force-Distilled Vision-Language-Action Model for Contact-Rich Manipulation
FD-VLA:力感知的视觉-语言-动作模型用于接触密集的 manipulation
机构 * Advanced Robotics Centre, National University of Singapore(新加坡国立大学先进机器人中心) ; School of Electrical & Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院) ; College of Information Science and Technology, Eastern Institute of Technology(东部技术学院信息科学与技术学院) ; John A. Paulson School of Engineering and Applied Sciences, Harvard University(哈佛大学约翰·A·保罗森工程与应用科学学院) ; Advanced Robotics Centre at National University of Singapore(新加坡国立大学先进机器人中心) ; Singapore Institute of Manufacturing Technology, Agency for Science, Technology and Research (A*STAR)(新加坡制造技术研究所,科技研究局(A*STAR))
专题命中 GUI与屏幕智能体 :VLM(abstract);分类 cs.CV
AI总结 本文提出FD-VLA模型,通过力蒸馏模块在不依赖物理力传感器的情况下实现接触密集任务的力感知,提升机器人视觉-语言-动作的鲁棒性与实用性。
Comments ICRA 2026 Accepted