3D Modality-Aware Pre-training for Vision-Language Model in MRI Multi-organ Abnormality Detection
面向MRI多器官异常检测的3D模态感知预训练
Haowen Zhu, Ning Yin, Xiaogen Zhou
机构
*
School of Electronic, Electrical Engineering and Physics, Fujian University of Technology(福建工程学院电子电气工程学院)
;
School of Computer Science and Engineering, Southeast University, China(东南大学计算机科学与工程学院)
;
Department of Medical Imaging, Suzhou Traditional Chinese Medicine Hospital, China(苏州中医医院影像科)
Video TokenCom: Textual Intent-Guided Multi-Rate Video Token Communications with UEP-Based Adaptive Source-Channel Coding
视频令牌通信:基于UEP的自适应源信道编码的文本意图引导多速率视频令牌通信
Jingxuan Men, Mahdi Boloursaz Mashhadi, Ning Wang, Yi Ma, Mike Nilsson, Rahim Tafazolli
机构
*
GIC & 6GIC, Institute for Communication Systems (ICS), University of Surrey(5GIC与6GIC,通信系统研究所(ICS),塞夫尔大学)
;
Smart Internet Lab, University of Bristol(智能互联网实验室,布里斯托尔大学)
;
British Telecom Research Lab, BT Group plc(英国电信研究实验室,BT集团)
专题命中
VLM训练与架构
:vision-language model(abstract);multimodal large language model(abstract);分类 cs.LG