Comments6 pages, 8 figures. Presented at 2025 IEEE International Workshop on Machine Learning for Signal Processing (MLSP), August 31 - September 3, 2025, Istanbul, Turkey
CommentsThe first two authors contributed equally. This work has been accepted to the Neural Information Processing Systems (NeurIPS) 2025 Datasets & Benchmark Track. Project Page: https://rf100-vl.org/
RadVLM: A Multitask Conversational Vision-Language Model for Radiology
Nicolas Deperrois, Hidetoshi Matsuo, Samuel Ruipérez-Campillo, Moritz Vandenhirtz, Sonia Laguna, Alain Ryser, Koji Fujimoto, Mizuho Nishio, Thomas M. Sutter, Julia E. Vogt, Jonas Kluckert, Thomas Frauenfelder, Christian Blüthgen, Farhad Nooralahzadeh, Michael Krauthammer
机构
*
Department of Radiology, Kobe University(金泽大学放射科)
;
Department of Computer Science, ETH Zurich(苏黎世联邦理工学院计算机科学系)
;
Department of Advanced Imaging in Medical Magnetic Resonance, Kyoto University(京都大学医学磁共振高级成像部门)
;
Department of Quantitative Biomedicine, University of Zurich(苏黎世大学定量生物医学系)
;
Diagnostic and Interventional Radiology, University Hospital Zurich(苏黎世大学医院诊断与介入放射科)
An Empirical Analysis of VLM-based OOD Detection: Mechanisms, Advantages, and Sensitivity
Yuxiao Lee, Xiaofeng Cao, Wei Ye, Jiangchao Yao, Jingkuan Song, Heng Tao Shen
机构
*
School of Artificial Intelligence, Jilin University(吉林大学人工智能学院)
;
School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院)
;
College of Electronic and Information Engineering, Tongji University(同济大学电子与信息工程学院)
;
CMIC, Shanghai Jiao Tong University(上海交通大学CMIC)
;
Engineering Research Center of Intelligent Finance, Ministry of Education(教育部智能金融工程研究中心)
;
Center for Future Media, University of Electronic Science and Technology of China(电子科技大学未来媒体中心)
Seeing the Trees for the Forest: Rethinking Weakly-Supervised Medical Visual Grounding
Ta Duc Huy, Duy Anh Huynh, Yutong Xie, Yuankai Qi, Qi Chen, Phi Le Nguyen, Sen Kim Tran, Son Lam Phung, Anton van den Hengel, Zhibin Liao, Minh-Son To, Johan W. Verjans, Vu Minh Hieu Phan
机构
*
Australian Institute for Machine Learning, University of Adelaide(澳大利亚机器学习研究所,阿德莱德大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
Macquarie University(麦考瑞大学)
;
Hanoi University of Science and Technology(河内科学技术大学)
;
University of Wollongong(沃林根大学)
;
Flinders University(弗林德斯大学)
机构
*
Xi’an Jiaotong University(西安交通大学)
;
ARC Lab, Tencent PCG(腾讯PCG ARC实验室)
;
City University of Hongkong(香港城市大学)
;
Institute of Automation, CAS(中国科学院自动化研究所)
;
Harvard University(哈佛大学)
;
vivo Mobile Communication Co.(vivo移动通信公司)
专题命中
视觉定位与Grounding
:grounding(title,abstract);multimodal large language model(abstract);分类 cs.CV、cs.AI
MedGround-R1: Advancing Medical Image Grounding via Spatial-Semantic Rewarded Group Relative Policy Optimization
Huihui Xu, Yuanpeng Nie, Hualiang Wang, Ying Chen, Wei Li, Junzhi Ning, Lihao Liu, Hongqiu Wang, Lei Zhu, Jiyao Liu, Xiaomeng Li, Junjun He
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Shanghai Innovation Institute(上海创新研究院)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
Fudan University(复旦大学)
Efficient Adaptation For Remote Sensing Visual Grounding
Hasan Moughnieh, Mohamad Chalhoub, Hasan Nasrallah, Cristiano Nattero, Paolo Campanella, Giovanni Nico, Ali J. Ghandour
机构
*
American University of Beirut(贝鲁特美国大学)
;
Lebanese University(黎巴嫩大学)
;
RASID SARL
;
WASDI
;
Institute for Applied Mathematics, CNR Bari(应用数学研究所,CNR巴里)
;
National Center for Remote Sensing, CNRS(遥感国家中心,CNRS)
Beyond Bare Queries: Open-Vocabulary Object Grounding with 3D Scene Graph
Sergey Linok, Tatiana Zemskova, Svetlana Ladanova, Roman Titkov, Dmitry Yudin, Maxim Monastyrny, Aleksei Valenkov
机构
*
Center for Cognitive Modeling, Moscow Institute of Physics and Technology(认知建模中心,莫斯科物理技术学院)
;
AIRI
;
Sberbank of Russia, Robotics Center(俄罗斯储蓄银行机器人中心)