机构
*
School of Software Engineering, South China University of Technology(软件工程学院,华南理工大学)
;
School of Future Technology, South China University of Technology(未来技术学院,华南理工大学)
;
Shien-Ming Wu School of Intelligent Engineering, South China University of Technology(智能工程学院,华南理工大学)
机构
*
University of Chinese Academy of Sciences (UCAS)(中国科学院大学)
;
New Laboratory of Pattern Recognition (NLPR), CASIA(中国科学院模式识别新技术实验室)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(中国科学院多模态人工智能系统国家重点实验室)
;
Hong Kong Institute of Science & Innovation, CASIA(香港科学与创新研究院)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Tianjin University(天津大学)
;
South China Hospital, Medical School, Shenzhen University(深圳大学医学院南方医院)
专题命中
视觉定位与Grounding
:grounding(title,abstract);multimodal large language model(abstract);MLLM(abstract);分类 cs.CV
机构
*
The University of Tokyo(东京大学)
;
S-Lab, Nanyang Technological University(南洋理工大学S实验室)
;
Duke University(杜克大学)
;
Salesforce AI Research(Salesforce人工智能研究)
;
Nanyang Technological University(南洋理工大学)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
Tokyo University of Science(东京科学大学)
专题命中
视觉定位与Grounding
:vision language model(title,abstract);VLM(abstract,comments);分类 cs.CV、cs.AI、cs.LG