MiniGPT-Pancreas: Multimodal Large Language Model for Pancreas Cancer Classification and Detection
Andrea Moglia, Elia Clement Nastasio, Luca Mainardi, Pietro Cerveri
机构
*
Department of Electronics, Information, and Bioengineering(电子、信息与生物工程系)
;
Polytechnic University of Milan(米兰理工学院)
;
Department of Industrial, and Information Engineering(工业与信息工程系)
;
University of Pavia(帕维亚大学)
专题命中
视觉定位与Grounding
:multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Journal refMoglia, A., Nastasio, E.C., Mainardi, L. et al. MiniGPT-Pancreas: Multimodal Large Language Model for Pancreas Cancer Observation and Localization in CT Images. J Healthc Inform Res (2025)
CommentsAccepted by ECCV 2024. A comprehensive and hierarchical 3D reasoning grounding benchmark in the era of foundation models. Project page: https://zcmax.github.io/projects/ScanReason
GridVAD: Open-Set Video Anomaly Detection via Spatial Reasoning over Stratified Frame Grids
GridVAD: 通过分层帧网格上的空间推理实现开放集视频异常检测
Mohamed Eltahir, Ahmed O. Ibrahim, Obada Siralkhatim, Tabarak Abdallah, Sondos Mohamed
机构
*
King Abdullah University of Science and Technology (KAUST)(阿卜杜拉国王科技大学)
;
Independent Researcher(独立研究员)
;
National Center for Research (NCR)(国家研究中心)
VLC Fusion: Vision-Language Conditioned Sensor Fusion for Robust Object Detection
VLC Fusion:面向鲁棒目标检测的视觉-语言条件传感器融合
Aditya Taparia, Noel Ngu, Mario Leiva, Joshua Shay Kricheli, John Corcoran, Nathaniel D. Bastian, Gerardo Simari, Paulo Shakarian, Ransalu Senanayake
机构
*
Arizona State University(亚利桑那州立大学)
;
Department of Computer Science and Engineering, Universidad Nacional del Sur and Institute for Computer Science and Engineering(计算机科学与工程系,国家南方大学和计算机科学与工程研究所)
;
U.S. Department of Defense(美国国防部)
;
United States Military Academy(美国军事学院)
;
Syracuse University(雪城大学)
Hallucination Detection and Correction in Medical VLMs via Counter-Evidence Verification
基于反事实证据验证的医学视觉语言模型幻觉检测与纠正
Nan Zhou, Ke Zou, Meng Liu, Linchao He, Jiaqi Zhu, Yi Zhang, Hu Chen, Huazhu Fu
机构
*
College of Computer Science, Sichuan University(四川大学计算机科学学院)
;
Yong Loo Lin School of Medicine, National University of Singapore(新加坡国立大学杨潞龄医学院)
;
Key Laboratory of Data Protection and Intelligent Management, Ministry of Education, Sichuan University(四川大学数据保护与智能管理教育部重点实验室)
;
National Key Laboratory of Autonomous Intelligent Unmanned Systems, Beijing Institute of Technology(北京理工大学自主智能无人系统国家重点实验室)
;
Institute of High Performance Computing (IHPC), Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局高性能计算研究所)
机构
*
Laboratory of IEMN, CNRS, Centrale Lille, UMR 8520, Univ. Polytechnique Hauts-de-France(伊姆纳实验室,国家科学研究中心,里尔中央理工大学,UMR 8520,法国高等技术大学)
;
Khalifa University(卡利法大学)
;
School of Communication and Information Engineering, Shanghai University(上海大学通信与信息工程学院)
;
Sorbonne Center for Artificial Intelligence, Sorbonne University Abu Dhabi(索邦人工智能中心,索邦大学阿布扎克分校)
Stephanie Ng, CP Lim, SueJen Looi, Hendrik Zurlinden, David Nguyen, Lei Wei, Saeid Nahavandi, Hailing Zhou
机构
*
Swinburne University of Technology(斯winburne大学)
;
National Transport Research Organisation(国家交通运输研究组织)
;
Google Cloud(谷歌云)
;
Deakin University(德金大学)
机构
*
City University of Hong Kong(香港城市大学)
;
Shenzhen Loop Area Institute(深圳河套学院)
;
CAIR, HKISI, Chinese Academy of Sciences(中国科学院计算智能研究所)
;
UESTC(电子科技大学)
;
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs
UHR-Micro: 诊断和缓解地球观测VLMs中的分辨率错觉
Shuo Ni, Tong Wang, Jing Zhang, He Chen, Haonan Guo, Ning Zhang, Bo Du
机构
*
National Key Laboratory of Science and Technology on Space-Born Intelligent Information Processing(国家空间智能信息处理科技重点实验室)
;
Beijing Institute of Technology(北京理工大学)
;
School of Computer Science(计算机学院)
;
Wuhan University(武汉大学)
;
Zhongguancun Academy(中关村学院)
;
State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing(测绘遥感信息工程国家重点实验室)
;
Hong Kong Polytechnic University(香港理工大学)
SoccerLens: Grounded Soccer Video Understanding Beyond Accuracy
SoccerLens: 超越准确性的 grounded 球赛视频理解
Ismael Elsharkawi, Ahmed Sait, Silvio Giancola, Bernard Ghanem, Hossam Sharara, Abdelrahman Eldesokey
机构
*
Department of Computer Science and Engineering, The American University in Cairo(美国亚历山大大学计算机科学与工程系)
;
Image And Visual Understanding Lab (IVUL), KAUST(卡塔尔大学图像与视觉理解实验室)