Ego-Grounding for Personalized Question-Answering in Egocentric Videos
面向第一视角视频的个性化问答中的自我定位
机构 * University of Science and Technology of China(中国科学技术大学) ; National University of Singapore(新加坡国立大学)
专题命中 视觉问答 :grounding(title,abstract);multimodal large language model(abstract);分类 cs.CV、cs.AI
AI总结 本文提出MyEgo数据集,用于评估多模态大语言模型在需要自我定位的个性化问答中的能力,发现现有模型在自我记忆和推理方面存在显著不足。
Comments To appear at CVPR'26