Group3D: MLLM-Driven Semantic Grouping for Open-Vocabulary 3D Object Detection
Group3D: 基于多模态大语言模型的开放词汇3D目标检测语义分组
机构 * Department of Artificial Intelligence, Sungkyunkwan University(成均馆大学人工智能系) ; Department of Artificial Intelligence, Yonsei University(延世大学人工智能系)
专题命中 其他多模态 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV
AI总结 Group3D通过将语义约束整合到实例构建过程中,解决多视角开放词汇3D目标检测中的几何过合并问题,实现语义与几何一致性融合的检测框架。
Comments 24 pages, 7 figures, Project page: https://ubin108.github.io/Group3D/