TOLiD: Bridging the Architecture Gap in Vision Foundation Model to LiDAR Pretraining via Token Lifting for Distillation
TOLiD:通过用于蒸馏的令牌提升弥合视觉基础模型与激光雷达预训练之间的架构差距
机构 * SAIVT Group, School of Electrical Engineering and Robotics, Queensland University of Technology(澳大利亚昆士兰科技大学电气工程与机器人学院SAIVT组) ; CSIRO Robotics, CSIRO(澳大利亚联邦科学与工业研究组织机器人部)
专题命中 激光雷达 :LiDAR(title,abstract);分类 cs.RO、cs.CV
AI总结 研究提出TOLiD方法,通过耦合激光雷达主干与从冻结VFM教师初始化的学生ViT,利用视锥体池化、注意力及掩码双线性采样等技术,解决模态和架构差距问题,在多数据集上评估显示能提升迁移效果。
Comments Accepted to The IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2026