Improving Adversarial Transferability on Vision-Language Pre-training Models via Surrogate-Specific Bias Correction
通过代理特定偏差校正提高视觉-语言预训练模型上的对抗迁移性
Lijia Yu, Jiuxin Cao, Yuchen Qiang, Changhao Chen, Yifei Huang, Bo Liu
机构
*
School of Cyber Science and Engineering, Southeast University(东南大学网络空间安全学院)
;
Purple Mountain Laboratories(紫金山实验室)
;
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
CommentsThis is the author version prior to incorporating the camera-ready comments. The final version will be included in the Proceedings of the European Conference on Computer Vision (ECCV) 2026
机构
*
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
Center for Frontier AI Research, Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局前沿人工智能研究中心)
;
School of Electrical Engineering and Automation, Fuzhou University(福州大学电气工程与自动化学院)
;
ByteDance(字节跳动)
Memory-Supported Synergistic Adaptation for Training-Free Test-Time Medical Image Segmentation
用于无训练测试时医学图像分割的内存支持协同适应
Lingrui Li, Nan Pu, Dong Zhao, Wenjing Li, Andrew P French, Zhun Zhong, Xin Chen
机构
*
School of Computer Science, University of Nottingham(诺丁汉大学计算机科学学院)
;
School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院)
;
Information Systems Technology and Design Pillar, Singapore University of Technology and Design(新加坡科技设计大学信息系统技术与设计支柱)
ESC: Emotional Self-Correction for Reliable Vision-Language Models
ESC:面向可靠视觉语言模型的情感自我纠正
Tien-Huy Nguyen, Minh-Nhat Nguyen, Nguyen Nhat Huy, Hung Viet Nguyen, Huy Nguyen Minh Nhat, Thanh-Huy Nguyen, Cuong Tuan Nguyen, Hoang M. Le, Dat Nguyen, Phat Kim Huynh, Min Xu, Ulas Bagci
机构
*
1 GenAI4E Lab 2 University of Information Technology, Ho Chi Minh City, Vietnam 3 Universit\"at Trier, Germany 4 Ho Chi Minh University of Technology, Ho Chi Minh City, Vietnam 5 PAMI Lab, Vietnamese German University, Vietnam 6 Vietnam National University, Ho Chi Minh City, Vietnam 7 Carnegie Mellon University, USA 8 Omoshiroi AI, USA 9 Harvard University, USA 10 Basis Research Institute 11 PASSIO Laboratory, North Carolina A\&T State University, USA 12 Mohamed bin Zayed University of Artificial Intelligence, UAE 13 Northwestern University, USA [4pt] Equal contribution. Corresponding author
Revealing Training Data Exposure in Vision Language Large Models via Parameter Gradients
揭示视觉语言大模型中的训练数据暴露:基于参数梯度的方法
Zhihao Zhu, Hongyi Tang, Yi Yang, Ahmed Abbasi
机构
*
Department of Information Systems, Business Statistics and Operations Management (ISOM), Hong Kong University of Science and Technology, Hong Kong, China(信息系统、商业统计与运营管理系(ISOM),香港科技大学,香港,中国)
;
Department of IT, Analytics, and Operations, University of Notre Dame, Notre Dame, Indiana, USA(信息技术、分析与运营系,诺丁汉大学,诺丁汉,印第安纳州,美国)
机构
*
University of Virginia(弗吉尼亚大学)
;
J. Crayton Pruitt Family Department of Biomedical Engineering, Herbert Wertheim College of Engineering, University of Florida(佛罗里达大学赫伯特·韦特海姆工程学院J. Crayton Pruitt家庭生物医学工程系)
Adjudicated Captioning: Multi-Agent Alignment Scoring and Consensus-Distilled Beam Arbitration for Strict Zero-Shot Image Captioning
裁决式字幕生成:用于严格零样本图像字幕生成的多智能体对齐评分与共识蒸馏束仲裁
Duy Tran Thanh, Thien-Phuc Doan, Long Nguyen-Vu, Ngo Tan Vu Khanh
机构
*
AI Platform OneNexus, OneMount(OneMount AI平台OneNexus)
;
School of Electronic Engineering, Soongsil University(崇实大学电子工程学院)
;
MoAdata(MoAdata公司)
;
University of Economics Ho Chi Minh City (UEH)(胡志明市经济大学)
Comments12 pages, 13 figures. Accepted at EXPLIMED 2026 (Third Workshop on Explainable Artificial Intelligence for the medical domain), IJCAI-ECAI 2026
Self-supervision drives representational convergence in medical foundation models more than clinical supervision
自我监督比临床监督更能推动医学基础模型中的表征趋同
Soroosh Tayebi Arasteh, Sebastian Ziegelmayer, Mahshad Lotfinia, Lisa Adams, Sven Nebelung, Jakob Nikolas Kather, Daniel Truhn
机构
*
RWTH Aachen University(亚琛工业大学)
;
University Hospital RWTH Aachen(亚琛工业大学附属医院)
;
Technical University of Munich(慕尼黑工业大学)
;
Friedrich-Alexander-Universität Erlangen-Nürnberg(埃尔朗根-纽伦堡弗里德里希-亚历山大大学)
;
Technical University Dresden(德累斯顿工业大学)
;
University Hospital Dresden(德累斯顿大学附属医院)
;
University Hospital Heidelberg(海德堡大学附属医院)
MEDLAYXPLAIN: Benchmarking the Expert-Lay Gap in Medical Vision-Language Models
MEDLAYXPLAIN: 医学视觉语言模型中的专家-外行差距基准测试
Han Jang, Junhyeok Lee, Songsoo Kim, Chae Young Lim, Hyeonjin Goh, Heeseong Eum, Kyu Sung Choi
机构
*
Seoul National University(首尔大学)
;
Seoul National University Hospital(首尔大学医院)
;
Seoul National University College of Medicine(首尔大学医学院)
;
Sungkyunkwan University School of Medicine(成均馆大学医学院)
Vision-language models for chest radiography do not always need the image
胸部X光片的视觉-语言模型并不总是需要图像
Mahshad Lotfinia, Sebastian Ziegelmayer, Lisa Adams, Daniel Truhn, Andreas Maier, Soroosh Tayebi Arasteh
机构
*
Pattern Recognition Lab, Friedrich-Alexander-Universität Erlangen-Nürnberg(弗里德里希-亚历山大-埃尔朗根-纽伦堡大学模式识别实验室)
;
Department of Diagnostic and Interventional Radiology, TUM University Clinic, School of Medicine and Health, Klinikum rechts der Isar, Technical University of Munich(慕尼黑工业大学医学院与健康学院伊萨尔河右岸医院诊断与介入放射学系)
;
Lab for AI in Medicine, RWTH Aachen University(亚琛工业大学医学人工智能实验室)
;
Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen(亚琛工业大学医院诊断与介入放射学系)
TEVI: Text-Conditioned Editing of Visual Representations via Sparse Autoencoders for Improved Vision-Language Alignment
TEVI: 基于稀疏自编码器的文本条件视觉表示编辑以改进视觉-语言对齐
Sweta Mahajan, Sukrut Rao, Jiahao Xie, Alexander Koller, Bernt Schiele
机构
*
Max Planck Institute for Informatics, Saarland Informatics Campus, Saarbrücken, Germany(马克斯·普朗克研究所信息学院,萨尔兰信息学院,德国萨尔布吕肯)
;
Department of Language Science and Technology, Saarland University, Saarbrücken, Germany(语言科学与技术系,萨尔兰大学,德国萨尔布吕肯)