CommentsAccepted version. Revised to match the version accepted to the 2026 IEEE Symposium on Security and Privacy (SP); added publication information and DOI
Journal refProceedings of the 2026 IEEE Symposium on Security and Privacy (SP), pp. 98-117, IEEE Computer Society, 2026
Bridging Mechanistic Interpretability and Prompt Engineering with Gradient Ascent for Interpretable Persona Control
通过梯度上升实现可解释人格控制与提示工程的桥梁
Harshvardhan Saini, Yiming Tang, Dianbo Liu
机构
*
Department of Computer Science(计算机科学系)
;
Indian Institute of Technology (ISM), Dhanbad(印度理工学院(ISM),丹巴德)
;
National University of Singapore(新加坡国立大学)
;
CIFAR Fellow(CIFAR研究员)
机构
*
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
;
Beijing Academy of Artificial Intelligence (BAAI)(北京智源人工智能研究院)
;
Beihang University(北京航空航天大学)
;
Eastern Institute of Technology, Ningbo(宁波东方理工大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Microsoft Research Asia (MSRA)(微软亚洲研究院)
Trustworthy Predictive Distributions for Tail Events with Semiparametric Diagnostic Transport Maps
面向尾部事件的可信预测分布:基于半参数诊断传输图
Elizabeth Cucuzzella, Rafael Izbicki, Ann B. Lee
机构
*
Department of Statistics and Data Science, Carnegie Mellon University(统计与数据科学系,卡内基梅隆大学)
;
Department of Statistics, Federal University of Sao Carlos(统计系,圣卡洛斯联邦大学)
TextDS: Parameter-Efficient Representation Alignment for Scene Text Detection under Distribution Shifts
TextDS: 分布偏移下场景文本检测的参数高效表示对齐
Boyuan Chen, Zichen Dang, Chuang Yang, Lap-Pui Chau, Yi Wang
机构
*
School of Electrical Engineering, Xi’an Jiaotong University(西安交通大学电气工程学院)
;
Department of Electrical and Electronic Engineering, The Hong Kong Polytechnic University(香港理工大学电机及电子工程学系)