Synthetic Vasculature and Pathology Enhance Vision-Language Model Reasoning
合成血管和病理增强视觉-语言模型推理
Chenjun Li, Cheng Wan, Laurin Lux, Alexander Berger, Richard B. Rosen, Martin J. Menten, Johannes C. Paetzold
机构
*
Cornell University(康奈尔大学)
;
Weill Cornell Medicine(韦尔·康奈尔医学)
;
Technical University of Munich(慕尼黑技术大学)
;
New York Eye and Ear Infirmary of Mount Sinai(圣文森特医院)
;
Cornell Tech(康奈尔科技)
Vision-Language Models for Automated 3D PET/CT Report Generation
用于自动化3D PET/CT报告生成的视觉-语言模型
Wenpei Jiao, Kun Shang, Hui Li, Ke Yan, Jiajin Zhang, Guangjie Yang, Lijuan Guo, Yan Wan, Xing Yang, Dakai Jin, Zhaoheng Xie
机构
*
Institute of Medical Technology and National Biomedical Imaging Center, Peking University(北京大学医学技术研究院和国家生物医学成像中心)
;
Peking University People’s Hospital(北京大学人民医院)
;
Peking University Third Hospital(北京大学第三医院)
;
DAMO Academy, Alibaba Group(阿里巴巴集团达摩院)
;
The Affiliated Hospital of Qingdao University(青岛大学附属医院)
;
The First Affiliated Hospital of Henan Medical University(河南医科大学第一附属医院)
;
Jiujiang City Key Laboratory of Cell Therapy, Jiu Jiang NO.1 People’s Hospital(九江市细胞治疗重点实验室,九江市第一人民医院)
Abn-BLIP: Abnormality-aligned Bootstrapping Language-Image Pre-training for Pulmonary Embolism Diagnosis and Report Generation from CTPA
Zhusi Zhong, Yuli Wang, Lulu Bi, Zhuoqi Ma, Sun Ho Ahn, Christopher J. Mullin, Colin F. Greineder, Michael K. Atalay, Scott Collins, Grayson L. Baird, Cheng Ting Lin, Webster Stayman, Todd M. Kolb, Ihab Kamel, Harrison X. Bai, Zhicheng Jiao
机构
*
Department of Diagnostic Imaging, Brown University Health(布朗大学健康中心诊断影像科)
;
Warren Alpert Medical School of Brown University(布朗大学沃伦·阿尔珀特医学院)
;
Department of Biomedical Engineering, Johns Hopkins University School of Medicine(约翰霍普金斯大学医学院生物医学工程系)
;
Department of Radiology and Radiological Sciences, Johns Hopkins University School of Medicine(约翰霍普金斯大学医学院放射科)
;
Johns Hopkins University Division of Pulmonary and Critical Care Medicine(约翰霍普金斯大学肺科与重症医学科)
;
Department of Radiology, University of Colorado School of Medicine(科罗拉多大学医学院放射科)
CommentsUpon further review, we identified that our dataset requires optimization to ensure research reliability and accuracy. Additionally, considering the target journal's latest submission policies, we believe comprehensive manuscript revisions are necessary
Point, Detect, Count: Multi-Task Medical Image Understanding with Instruction-Tuned Vision-Language Models
Sushant Gautam, Michael A. Riegler, Pål Halvorsen
机构
*
Simula Metropolitan Center for Digital Engineering (SimulaMet), Norway(Simula数字工程中心(SimulaMet))
;
Oslo Metropolitan University (OsloMet), Norway(奥斯陆 Metropolitan 大学(OsloMet))
;
Simula Research Laboratory, Norway(Simula研究实验室)
机构
*
School of Biomedical Engineering, Tsinghua University(清华大学生物医学工程学院)
;
School of Biomedical Engineering, Shanghai Jiao Tong University(上海交通大学生物医学工程学院)
;
DAMO Academy, Alibaba Group(阿里云达摩院)
;
Hupan Laboratory(壶辰实验室)
;
Department of Biomedical Engineering, National University of Singapore(新加坡国立大学生物医学工程系)
;
Department of Radiology, Guizhou Provincial People’s Hospital(贵州省级人民医院放射科)
;
Department of Radiology, The First Affiliated Hospital, Zhejiang University School of Medicine(浙江大学医学院附属第一医院放射科)
;
Department of Radiology, Shanghai Sixth People’s Hospital Affiliated to Shanghai Jiao Tong University School of Medicine(上海交通大学医学院附属第六人民医院放射科)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)