Multi-Modal Cross-Domain Alignment Network for Video Moment Retrieval
多模态跨域对齐网络用于视频时刻检索
机构 * Hubei Key Laboratory of Distributed System Security(湖北分布式系统安全重点实验室) ; Hubei Engineering Research Center on Big Data Security(湖北大数据安全工程研究中心) ; School of Cyber Science and Engineering(网络安全学院) ; Huazhong University of Science and Technology(华中科技大学) ; Wangxuan Institute of Computer Technology(王轩计算机技术研究所) ; Peking University(北京大学) ; School of Computer Science and Technology(计算机科学与技术学院) ; Key Laboratory of Information Storage System Ministry of Education of China(信息存储系统教育部重点实验室)
专题命中 跨模态检索 :multi-modal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI、cs.MM
AI总结 提出多模态跨域对齐网络,通过域对齐、跨模态对齐和特定对齐三个模块,解决跨域视频时刻检索中域差异和语义鸿沟问题。
Comments Accepted by IEEE Transactions on Multimedia