CMRAG: Co-modality-based visual document retrieval and question answering
CMRAG:基于共模的视觉文档检索与问答
机构 * Baidu Inc(百度公司) ; The University of Hong Kong(香港大学) ; Beihang University(北京航空航天大学) ; Peking University(北京大学)
专题命中 多模态RAG :retrieval-augmented generation(abstract);RAG(abstract);分类 cs.CL
AI总结 CMRAG通过统一编码模型和共模信息检索方法,提升多模态视觉文档问答系统的性能。
Comments Published at ICLR 2026 Workshop on Multimodal Intelligence