HAMMER: Harnessing MLLM via Cross-Modal Integration for Intention-Driven 3D Affordance Grounding
HAMMER: 通过跨模态整合利用大语言模型进行意图驱动的3D affordance grounding
机构 * The Hong Kong Polytechnic University(香港理工大学) ; Huazhong University of Science and Technology(华中科技大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);MLLM(title,abstract);multimodal large language model(abstract);分类 cs.CV
AI总结 HAMMER通过跨模态整合多模态大语言模型,实现意图驱动的3D affordance grounding,提升3D表示的准确性和鲁棒性。
Comments Accepted by CVPR 2026. Project Page: https://rayyoh.github.io/Hammer