机构
*
School of Computer Science, Peking University(北京大学计算机科学学院)
;
School of Computer Science, Zhejiang University(浙江大学计算机科学学院)
;
School of Cyber Science and Engineering, Wuhan University(武汉大学网络科学与工程学院)
;
School of Information Science and Technology, Tibet University(西藏大学信息科学与技术学院)
Multi-Modal Object Re-Identification with Dual Semantic Guidance and Global-Local Mutual Modulation
融合双语义引导与全局-局部互调制的多模态物体重识别
Weixiang Zhou, Xingguo Xu, Yuhao Wang, Cong Wang, Yang Yang, Zhixun Su, Jinshan Pan
机构
*
Dalian University of Technology(大连理工大学)
;
University of California, San Francisco(加利福尼亚大学旧金山分校)
;
Nanjing University of Science and Technology(南京理工大学)
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
Sun Yat-sen University(中山大学)
;
InfiX.ai
;
PolyU-Daya Bay Technology and Innovation Research Institute(PolyU-大亚湾技术与创新研究院)
CHARM: A Multimodal Graph Foundation Model with Hierarchical Context Modeling for Zero-Shot Transfer
CHARM:一种用于零样本迁移的具有分层上下文建模的多模态图基础模型
Ankang Yang, Jitao Zhao, Di Jin, Yuxiao Huang, Dongxiao He
机构
*
School of Computer Science and Technology, College of Intelligence and Computing, Tianjin University(天津大学智能与计算学部计算机科学与技术学院)
;
Data Science Program, Columbian College of Arts and Sciences, The George Washington University(乔治华盛顿大学文理学院哥伦比亚艺术与科学学院数据科学项目)
QATMA: Quantization-Aware Training with Multimodal Alignment for Open-Vocabulary Object Detection
CR-QAT: 基于课程关系量化感知训练的开放词汇目标检测
Jinyeong Park, Donghwa Kang, Seunghwan An, Insoo Kim, Brent ByungHoon Kang, Hyeongboo Baek, Jibum Kim
机构
*
Incheon National University, Incheon, South Korea(韩国仁川国立大学)
;
KAIST, Daejeon, South Korea(韩国成均馆大学)
;
University of Seoul, Seoul, South Korea(首尔大学)
机构
*
Indian Institute of Technology Ropar(印度理工学院罗帕尔分校)
;
RoentGen Health(伦琴健康公司)
;
Deakin University(迪肯大学)
;
University of Central Florida(中佛罗里达大学)
;
Queensland University of Technology(昆士兰科技大学)
;
Murdoch University(莫道克大学)
机构
*
School of Artificial Intelligence, Beihang University, Beijing, China(北京航空航天大学人工智能学院)
;
School of Computer Science and Engineering, Beihang University, Beijing, China(北京航空航天大学计算机科学与工程学院)
Blurring Modal Boundaries: A Unified Survey from Single- to Multi-Modal Person Re-ldentification
模糊模态边界:从单模态到多模态行人重识别的统一综述
Xiao Wang, Bing Wang, Bin Yang, Cuiqun Chen, Xin Xu, Mang Ye
机构
*
Wuhan University of Science and Technology(武汉科技大学)
;
Hubei Province Key Laboratory of Intelligent Information Processing and Real-time Industrial System(湖北省智能信息处理与实时工业系统重点实验室)
;
Wuhan University(武汉大学)
;
Anhui University(安徽大学)
Segregate, Refine, Integrate: Decomposing Multimodal Fusion for Sentiment Analysis
分离、细化、整合:用于情感分析的多模态融合分解
Alexios Filippakopoulos, Elias Kallioras, Nikolaos Xiros, Efthymios Georgiou, Alexandros Potamianos
机构
*
National Technical University of Athens(雅典国立技术大学)
;
Athena Research Center(雅典娜研究中心)
;
University of Bern(伯尔尼大学)
;
Archimedes AI(阿基米德人工智能公司)
;
Synaptic Bloom PBC(突触绽放有限责任公司)