GeoAware-VLA: Implicit Geometry Aware Vision-Language-Action Model
GeoAware-VLA: 基于隐式几何的视觉-语言-动作模型
机构 * Department of Robotics, Mohamed bin Zayed University of Artificial Intelligence(机器人系,Mohamed bin Zayed人工智能大学)
专题命中 VLA模型 :VLA(title,abstract);vision-language-action(title,abstract);action model(title);分类 cs.RO
AI总结 GeoAware-VLA通过整合几何先验提升视觉-语言-动作模型的视角不变性,显著提高零样本泛化能力,并在现实机器人平台中取得显著效果。
Comments Under Review, Project Page https://alisharey.github.io/GeoAware-VLA/