MARVL: Multi-Stage Guidance for Robotic Manipulation via Vision-Language Models
MARVL:通过视觉-语言模型实现机器人操作的多阶段引导
机构 * School of Intelligent Science and Technology, Nanjing University, China(南京大学智能科学与技术学院) ; National Key Laboratory for Novel Software Technology, School of Artificial Intelligence, Nanjing University, China(南京大学新型软件技术国家实验室,人工智能学院) ; School of Artificial Intelligence, Nanjing University, China(南京大学人工智能学院) ; MACS Lab, University of Washington(华盛顿大学MACS实验室)
专题命中 机器人操作 :manipulation(title,abstract);robotic(title,abstract);分类 cs.RO、cs.CV、cs.LG
AI总结 本文提出MARVL,通过视觉-语言模型实现机器人操作的多阶段引导,解决传统密集奖励函数设计中的问题,提升样本效率和鲁棒性。