机构
*
Department of Computer Science and Engineering, Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系)
;
Department of Pathology, Nanfang Hospital, Southern Medical University(南方医科大学南芳医院病理科)
;
Department of Pathology, School of Basic Medical Sciences, Southern Medical University(南方医科大学基础医学学院病理科)
;
Department of Anatomical and Cellular Pathology, Chinese University of Hong Kong(香港中文大学解剖与细胞病理学系)
;
Guangdong Provincial Key Laboratory of Molecular Tumor Pathology(广东省分子肿瘤病理学重点实验室)
;
Jinfeng Laboratory(锦风实验室)
;
Department of Chemical and Biological Engineering, Hong Kong University of Science and Technology(香港科技大学化学与生物工程系)
;
Division of Life Science, Hong Kong University of Science and Technology(香港科技大学生命科学系)
;
State Key Laboratory of Nervous System Disorders, The Hong Kong University of Science and Technology(香港科技大学神经系统疾病国家重点实验室)
;
HKUST Shenzhen-Hong Kong Collaborative Innovation Research Institute, The Hong Kong University of Science and Technology(香港科技大学深圳-香港协同创新研究院)
GazeXPErT: An Expert Eye-tracking Dataset for Interpretable and Explainable AI in Oncologic FDG-PET/CT Scans
GazeXPErT: 一种用于肿瘤FDG-PET/CT扫描可解释和可解释AI的专家眼动数据集
Joy T Wu, Daniel Beckmann, Sarah Miller, Alexander Lee, Elizabeth Theng, Stephan Altmayer, Ken Chang, David Kersting, Tomoaki Otani, Brittany Z Dashevsky, Hye Lim Park, Matteo Novello, Kip Guja, Curtis Langlotz, Ismini Lourentzou, Daniel Gruhl, Benjamin Risse, Guido A Davidzon
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Large Language Model Department, Tencent(腾讯大语言模型部)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Zhongguancun Academy(中关村学院)
专题命中
视觉定位与Grounding
:MLLM(title_cn,summary_cn);grounding(abstract,abstract_cn);multimodal large language model(abstract);分类 cs.CV
The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics
标准可解释模型:一种基于拉格朗日力学的可解释机器学习通用理论,用于演绎设计可解释方法
Pietro Barbiero, Giovanni De Felice, Mateo Espinosa Zarlenga, Francesco Giannini, Filippo Bonchi, Mateja Jamnik, Giuseppe Marra, Ruggero Noris
机构
*
IBM Research (CH)(IBM研究院(瑞士))
;
University of Oxford (UK)(牛津大学(英国))
;
University of Cambridge (UK)(剑桥大学(英国))
;
KU Leuven (BE)(鲁汶大学(比利时))
;
Institute of Physics of the Czech Academy of Sciences (CZ)(捷克科学院物理研究所(捷克))
机构
*
Peking University(北京大学)
;
Math Magic(数学魔术)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Guangdong Provincial Key Laboratory of Ultra High Definition Immersive Media Technology(广东省超高清沉浸媒体技术重点实验室)
;
Dalian University of Technology(大连理工大学)
专题命中
幻觉与鲁棒性
:vision-language model(title);VLM(abstract,abstract_cn);multimodal large language model(abstract);分类 cs.CV
CLAIR-Fin: An Adversarial Multi-Agent Framework for Claim-Level Verification and Adaptive Debate in Cross-Modal Financial QA
CLAIR-Fin:用于跨模态金融问答中声明级验证与自适应辩论的对抗性多智能体框架
Fatema Tuj Johora Faria, Mukaffi Bin Moin, Jubayer Al Mahmud, M. F. Mridha, Md. Alam Hossain
机构
*
Ahsanullah University of Science and Technology(阿萨努拉科技大学)
;
Jashore University of Science and Technology(杰索尔科技大学)
;
American International University - Bangladesh(孟加拉国美国国际大学)
Seeing is Free, Speaking is Not: Uncovering the True Energy Bottleneck in Edge VLM Inference
眼见无需耗能,言语却要代价:揭示边缘视觉语言模型推理中的真正能量瓶颈
Junfei Zhan, Haoxun Shen, Mingang Guo, Zixuan Huang, Tengjiao He
机构
*
University of Pennsylvania(宾夕法尼亚大学)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)
;
Jinan University(暨南大学)
Cytoarchitecture in Words: Weakly Supervised Vision-Language Modeling for Human Brain Microscopy
词中结构:用于人类大脑显微镜的弱监督视觉-语言建模
Matthew Sutton, Katrin Amunts, Timo Dickscheid, Christian Schiffer
机构
*
Institute of Neuroscience and Medicine (INM-1), Research Centre Jülich(神经科学与医学研究所(INM-1),焦耳研究中心)
;
Helmholtz AI, Research Centre Jülich(海德堡人工智能研究所,焦耳研究中心)
;
Institute for Brain Research, University Hospital Düsseldorf(脑研究所在杜塞尔多夫大学医院)
;
Computer Vision, Institute for Computational Visualistics, University of Koblenz(计算机视觉,计算视觉研究所,科布伦茨大学)
Lorenz Hufe, Niclas Griesshaber, Gavin Greif, Sebastian Oliver Eck, Pieter Francois, Wojciech Samek, Christian Schroeder de Witt, Philip Torr
机构
*
Torr Vision Group, Department of Engineering Science, University of Oxford(牛津大学工程科学系托尔视觉组)
;
Oxford Centre for Economic and Social History, University of Oxford(牛津大学牛津经济与社会史中心)
;
Faculty of Music, University of Oxford(牛津大学音乐学院)
;
Fraunhofer HHI(弗劳恩霍夫海因里希·赫兹研究所)
LLM-Based Generative Retrieval for Snapchat Content Recommendation
基于大语言模型的生成式检索用于Snapchat内容推荐
Liam Collins, Jiwen Ren, Donald Loveland, Bhuvesh Kumar, Clark Mingxuan Ju, Xuan Guo, Mo Li, Alvin Hou, Yi Cui, Peng Yang, Jian Wang, Saud Afzal Shafi, Nga Than, Ruiming Lu, Wenfeng Zhuo, Dongheng Li, Lili Zhang, Mingtao Zhang, Jinchao Ye, Vincent Xue, Chunhui Zhu, Neil Shah