Improving 3D Labeling in Self-Driving by Inferring Vehicle Information using Vision Language Models
通过利用视觉语言模型推断车辆信息以改进自动驾驶中的3D标注
机构 * Aurora Innovation, Inc.(Aurora创新公司)
专题命中 视觉定位与Grounding :vision language model(title,abstract);VLM(abstract,abstract_cn);分类 cs.CV
AI总结 本文提出了一种利用视觉语言模型推断车辆信息以提高自动驾驶中3D车辆标注精度的方法,通过零样本推理车辆信息,结合车辆型号和型号识别方法,提升了标注效率和质量。
Comments To appear in Proceedings of the IEEE Intelligent Vehicles Symposium (IV), 2026. Accepted for oral presentation