The Vision Encoder as a Privacy Boundary: Visual-Token Side Channels in Encoder-Free Vision-Language Models
视觉编码器作为隐私边界:无编码器视觉-语言模型中的视觉令牌侧信道
Chenyu Zhou, Qiliang Jiang, Shuning Wu, Xu Zhou
机构
*
School of Engineering, Institute of Science Tokyo(东京科学大学工学院)
;
College of Control Science and Engineering, Zhejiang University(浙江大学控制科学与工程学院)
;
Department of Electrical and Computer Engineering, National University of Singapore(新加坡国立大学电气与计算机工程系)
Comments17 pages, 5 main-paper figures. Accepted at the 34th ACM International Conference on Multimedia (ACM MM 2026). Includes the complete supplementary material
Comments12 pages, 13 figures. Accepted at EXPLIMED 2026 (Third Workshop on Explainable Artificial Intelligence for the medical domain), IJCAI-ECAI 2026
Continual Learning with Elastic Regularization and Synthetic Replay for Federated MLLM Fine-Tuning
用于联邦多模态大语言模型微调的弹性正则化和合成重放持续学习
Jing Liu, Chenxuanyin Zou, Jiayang Ren, Gaoyun Fang, Chengfang Li, Yan Wang, Zhenchao Ma, Bo Hu
机构
*
The University of British Columbia(英属哥伦比亚大学)
;
Fudan University(复旦大学)
;
Royal College of Science, Imperial College London(伦敦帝国理工学院皇家科学学院)
;
Dyson School of Design Engineering(戴森设计工程学院)
;
Suzhou Institute of Biomedical Engineering and Technology (SIBET), Chinese Academy of Sciences(中国科学院苏州生物医学工程技术研究所)
;
East China Normal University(华东师范大学)
专题命中
VLM训练与架构
:MLLM(title,abstract_cn);multimodal large language model(abstract);分类 cs.CV、cs.AI、cs.LG
CPS4: Class Prompt driven Semi-Supervised Spine Segmentation with Class-specific Consistency Constraint
CPS4: 基于类别提示的半监督脊柱分割与类别特定一致性约束
Qingtao Pan, Hongzan Sun, Bing Ji, Shuo Li
机构
*
School of Control Science and Engineering, Shandong University(山东大学控制科学与工程学院)
;
Department of Nuclear Medicine, Shengjing Hospital of China Medical University(中国医科大学附属盛京医院核医学科)
;
Department of Computer and Data Science, Case Western Reserve University(凯斯西储大学计算机与数据科学系)
;
Department of Biomedical Engineering, Case Western Reserve University(凯斯西储大学生物医学工程系)
专题命中
VLM训练与架构
:VLM(summary_cn,abstract);vision language model(abstract);分类 cs.CV
Small Vision-Language Models Know When They Are Wrong But Cannot Say So: A Two-Model Study of Stated versus Internal Confidence Under Realistic Image Degradation