Point2Act: Efficient 3D Distillation of Multimodal LLMs for Zero-Shot Context-Aware Grasping
Point2Act: 多模态大语言模型的高效3D蒸馏用于零样本情境感知抓取
机构 * Seoul National University(首尔国立大学) ; Massachusetts Institute of Technology(麻省理工学院)
专题命中 视频多模态 :multimodal(title,abstract);MLLM(abstract)
AI总结 Point2Act通过多模态大语言模型高效蒸馏实现零样本情境感知抓取,生成空间定位响应以支持实际操作任务。
Comments Accepted to ICRA 2026