DecAlign: Hierarchical Cross-Modal Alignment for Decoupled Multimodal Representation Learning
DecAlign:解耦多模态表示学习的分层跨模态对齐
机构 * Texas A&M University(德克萨斯A&M大学) ; University of Southern California(南加州大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(title,abstract);分类 cs.CV
AI总结 DecAlign通过分层跨模态对齐框架解耦多模态表示,结合原型引导最优传输策略和多模态Transformer提升语义一致性,实验显示在四个基准上优于现有方法。
Comments Accepted by ICLR 2026