机构
*
Dept. of Computer Science and Engineering, The Chinese University of Hong Kong, Hong Kong, China(计算机科学与工程系,香港中文大学,香港,中国)
;
Institute of Medical Intelligence and XR, The Chinese University of Hong Kong, Hong Kong, China(医学智能与XR研究所,香港中文大学,香港,中国)
机构
*
School of Intelligent Systems Engineering, Sun Yat-Sen University(中山大学智能系统工程学院)
;
DP Technology(DP技术公司)
;
Beijing University Of Posts and Telecommunications(北京邮电大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Tongji University(同济大学)
;
School of Integrated Innovation, Chulalongkorn University(朱拉隆功大学创新学院)
CUBIC: Concept Embeddings for Unsupervised Bias Identification using VLMs
David Méndez, Gianpaolo Bontempo, Elisa Ficarra, Roberto Confalonieri, Natalia Díaz-Rodríguez
机构
*
Dept. of Computer Science and Artificial Intelligence, DaSCI Institute, University of Granada(计算机科学与人工智能系,DaSCI研究所,格拉纳达大学)
;
Dept. of Engineering ”Enzo Ferrari”, University of Modena and Reggio Emilia(工程系,摩德纳和雷吉奥艾米利亚大学)
;
Dept. of Mathematics ’Tullio Levi-Civita’, University of Padova(数学系,帕多瓦大学)
专题命中
图文多模态
:image-text(abstract);分类 cs.CV、cs.AI
Comments8 pages, 3 figures, 5 tables. Accepted at IJCNN 2025; to appear in IEEE Xplore
Endo-CLIP: Progressive Self-Supervised Pre-training on Raw Colonoscopy Records
Yili He, Yan Zhu, Peiyao Fu, Ruijie Yang, Tianyi Chen, Zhihua Wang, Quanlin Li, Pinghong Zhou, Xian Yang, Shuo Wang
机构
*
Digital Medical Research Center, School of Basic Medical Sciences, Fudan University, Shanghai, China(上海复旦大学基础医学学院数字医学研究中心)
;
University College London, London, UK(伦敦大学学院)
;
Shanghai Key Laboratory of MICCAI, Shanghai, China(上海MICCAI重点实验室)
;
Endoscopy Center and Endoscopy Research Institute, Zhongshan Hospital, Fudan University, Shanghai, China(复旦大学中山医院内窥镜中心和内窥镜研究所)
;
Shanghai Collaborative Innovation Center of Endoscopy, Shanghai, China(上海内窥镜协同创新中心)
;
Shanghai Institute for Advanced Study of Zhejiang University, Shanghai, China(浙江大学上海高级研究院)
;
Alliance Manchester Business School, The University of Manchester, Manchester, UK(曼彻斯特大学曼彻斯特商业学校)
;
Data Science Institute, Imperial College London, London, UK(伦敦帝国理工学院数据科学研究院)
Reinforced Correlation Between Vision and Language for Precise Medical AI Assistant
Haonan Wang, Jiaji Mao, Lehan Wang, Qixiang Zhang, Marawan Elbatel, Yi Qin, Huijun Hu, Baoxun Li, Wenhui Deng, Weifeng Qin, Hongrui Li, Jialin Liang, Jun Shen, Xiaomeng Li
机构
*
Department of Electronic and Computer Engineering, HKUST(香港科技大学电子与计算机工程系)
;
Department of Radiology, Guangdong Provincial Key Laboratory of Malignant Tumor Epigenetics and Gene Regulation, Sun Yat-Sen Memorial Hospital, Sun Yat-Sen University(中山大学放射科、广东省恶性肿瘤表观遗传与基因调控重点实验室、中山纪念医院)
;
Department of Computer Science and Engineering, HKUST(香港科技大学计算机科学与工程系)
A Comprehensive Analysis for Visual Object Hallucination in Large Vision-Language Models
Liqiang Jing, Guiming Hardy Chen, Ehsan Aghazadeh, Xin Eric Wang, Xinya Du
机构
*
University of Texas at Dallas(德克萨斯大学达拉斯分校)
;
University of Massachusetts at Amherst(马萨诸塞大学阿姆赫斯特分校)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
Detecting and Understanding Hateful Contents in Memes Through Captioning and Visual Question-Answering
Ali Anaissi, Junaid Akram, Kunal Chaturvedi, Ali Braytee
机构
*
The University of Sydney, School of Computer Science(悉尼大学计算机科学学院)
;
University of Technology Sydney, School of Computer Science(新南威尔士大学技术学院)
;
University of Technology Sydney, TD School(新南威尔士大学TD学院)
;
Australian Catholic University, Peter Faber Business School(澳大利亚天主教大学彼得·法伯商学院)
专题命中
图文多模态
:multimodal(abstract);分类 cs.CV、cs.AI
Comments13 pages, 2 figures, 2025 International Conference on Computational Science
机构
*
University of Science and Technology of China(中国科学技术大学)
;
SKL of Processors, Institute of Computing Technology, CAS(中国科学院计算技术研究所处理器专项实验室)
;
National University of Singapore(新加坡国立大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)