Unlocking Multimodal Document Intelligence: From Current Triumphs to Future Frontiers of Visual Document Retrieval
解锁多模态文档智能:从当前成就到视觉文档检索的未来前沿
Yibo Yan, Jiahao Huo, Guanbo Feng, Mingdong Ou, Yi Cao, Xin Zou, Shuliang Liu, Yuanhuiyi Lyu, Yu Huang, Jungang Li, Kening Zheng, Xu Zheng, Philip S. Yu, James Kwok, Xuming Hu
机构
*
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Alibaba Cloud Computing(阿里云计算)
;
Hong Kong University of Science and Technology(香港科技大学)
;
University of Illinois Chicago(伊利诺伊大学芝加哥分校)
Asymptotically and Minimax Optimal Regret Bounds for Multi-Armed Bandits with Abstention
渐近最优和最小最大最优的多臂老虎机问题中的弃权后悔界
Junwen Yang, Tianyuan Jin, Vincent Y. F. Tan
机构
*
Institute of Operations Research and Analytics(运营研究与分析研究所)
;
National University of Singapore(新加坡国立大学)
;
Department of Electrical and Computer Engineering(电子与计算机工程系)
;
Department of Mathematics(数学系)
机构
*
School of Computer Science and Technology, UCAS(UCAS计算机科学与技术学院)
;
The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, CASIA(复杂系统认知与决策智能重点实验室)
;
School of Computer Science and Engineering, Beihang University(北航计算机科学与工程学院)
;
State Key Laboratory of Networking and Switching Technology, BUPT(网络与交换技术国家重点实验室)
;
AMAP, Alibaba Group(阿里妈妈实验室,阿里巴巴集团)
Cross-modal Fuzzy Alignment Network for Text-Aerial Person Retrieval and A Large-scale Benchmark
跨模态模糊对齐网络用于文本-空中人检索及一个大规模基准
Yifei Deng, Chenglong Li, Yuyang Zhang, Guyue Hu, Jin Tang
机构
*
State Key Laboratory of Opto-Electronic Information Acquisition and Protection Technology(光电信息采集与防护技术国家重点实验室)
;
School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院)
;
School of Artificial Intelligence, Anhui University(安徽大学人工智能学院)
;
The University of Hong Kong(香港大学)