机构
*
College of Computer Science and Artificial Intelligence, Fudan University, Shanghai, China(复旦大学计算机科学与人工智能学院,上海,中国)
;
Shanghai Key Laboratory of Intelligent Information Processing(上海智能信息处理重点实验室)
Intelligent Healthcare Imaging Platform: A VLM-Based Framework for Automated Medical Image Analysis and Clinical Report Generation
智能医疗影像平台:基于视觉语言模型的自动化医学图像分析与临床报告生成框架
Samer Al-Hamadani
机构
*
Automated Manufacturing Department/Al-Khwarizmi College of Engineering/ University of Baghdad/Gilgamesh University(自动化制造部门/阿尔·卡瓦尔齐米工程学院/巴格达大学/吉尔伽美什大学)
TTL: Test-time Textual Learning for OOD Detection with Pretrained Vision-Language Models
TTL: 用于基于预训练视觉-语言模型的分布外检测的测试时文本学习
Jinlun Ye, Jiang Liao, Runhe Lai, Xinhua Lu, Jiaxin Zhuang, Zhiyong Gan, Ruixuan Wang
机构
*
Sun Yat-sen University(中山大学)
;
China United Network Communications Corporation Limited Guangdong Branch(中国联合网络通信集团有限公司广东分公司)
;
Peng Cheng Laboratory(鹏城实验室)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Key Laboratory of Machine Intelligence and Advanced Computing, MOE(教育部机器智能与高级计算重点实验室)
机构
*
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
Huazhong University of Science and Technology(华中科技大学)
;
The University of Hong Kong(香港大学)
;
ByteDance Seed(字节跳动种子)
;
Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究院)
;
Shanghai Key Laboratory of Multimodal Embodied AI(上海多模态具身人工智能重点实验室)
Teaching Language Models Mechanistic Explainability Through MechSMILES
通过MechSMILES教会语言模型机制性可解释性
Théo A. Neukomm, Zlatko Jončev, Philippe Schwaller
机构
*
Laboratory of Artificial Chemical Intelligence (LIAC), Institut des Sciences et Ingénierie Chimiques, Ecole Polytechnique Fédérale de Lausanne (EPFL), Lausanne 1015, Switzerland(人工化学智能实验室(LIAC)、化学与工程学院、瑞士联邦理工学院(EPFL)、拉沃斯纳1015号)
机构
*
AMAP, Alibaba Group(阿里集团AMAP)
;
University of California at Merced(加州大学默塞德分校)
;
University of Queensland(昆士兰大学)
;
Case Western Reserve University(凯斯西储大学)
专题命中
视觉定位与Grounding
:multimodal large language model(abstract);分类 cs.CV