A Comprehensive Information-Decomposition Analysis of Large Vision-Language Models
大型视觉-语言模型的全面信息分解分析
机构 * The University of Tokyo(东京大学) ; Microsoft Research(微软研究院)
专题命中 音视频/视觉语言融合 :multimodal fusion(abstract);分类 cs.CV
AI总结 本文通过信息分解方法分析大型视觉-语言模型的信息谱,揭示任务模式和模型策略,为模型设计提供新视角。
Comments Accepted at ICLR 2026. Project page: https://riishin.github.io/pid-lvlm-iclr26/