ReasAlign: Reasoning Enhanced Safety Alignment against Prompt Injection Attack
ReasAlign: 基于推理增强的安全对齐以抵御提示注入攻击
Hao Li, Yankai Yang, G. Edward Suh, Ning Zhang, Chaowei Xiao
机构
*
Washington University in St. Louis(华盛顿大学圣路易斯分校)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
NVIDIA(NVIDIA公司)
;
Johns Hopkins University(约翰霍普金斯大学)
GuardNet: Ensemble Strategies of Shallow Neural Networks for Robust Prompt Injection and Jailbreak Detection
GuardNet: 用于鲁棒提示注入和越狱检测的浅层神经网络集成策略
Paulo Ricardo Ferreira Neves, Edson Rodrigues da Cruz Filho, Paulo Henrique Eleuterio Falsetti, João Vitor Pavan, Ian Degaspari, Henrique Vieira Laturrague, Patrick Vieira Laturrague, Guilherme Nielsen Dias, Marccello Wilson Perez Berto, Gustavo Voltani Von Atzingen
机构
*
Quickium Technology Ltd.(Quickium技术有限公司)
;
Federal University of São Carlos (UFSCar)(萨尔瓦多·卡罗斯联邦大学)
;
Federal Institute of Education, Science and Technology of São Paulo (IFSP)(圣保罗教育、科学和技术联邦研究所)
Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety
大规模安全:大型模型和智能体安全的全面综述
Xingjun Ma, Yifeng Gao, Yixu Wang, Ruofan Wang, Xin Wang, Ye Sun, Yifan Ding, Hengyuan Xu, Yunhao Chen, Yunhan Zhao, Hanxun Huang, Yige Li, Yutao Wu, Jiaming Zhang, Xiang Zheng, Yang Bai, Zuxuan Wu, Xipeng Qiu, Jingfeng Zhang, Yiming Li, Xudong Han, Haonan Li, Jun Sun, Cong Wang, Jindong Gu, Baoyuan Wu, Siheng Chen, Tianwei Zhang, Yang Liu, Mingming Gong, Tongliang Liu, Shirui Pan, Cihang Xie, Tianyu Pang, Yinpeng Dong, Ruoxi Jia, Yang Zhang, Shiqing Ma, Xiangyu Zhang, Neil Gong, Chaowei Xiao, Sarah Erfani, Tim Baldwin, Bo Li, Masashi Sugiyama, Dacheng Tao, James Bailey, Yu-Gang Jiang
机构
*
Fudan University(复旦大学)
;
The University of Melbourne(墨尔本大学)
;
Singapore Management University(新加坡国立大学)
;
Deakin University(德肯大学)
;
Hong Kong University of Science and Technology(香港科学与技术大学)
;
City University of Hong Kong(香港城市大学)
;
University of Oxford(牛津大学)
;
Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Shanghai Jiao Tong University(上海交通大学)
;
Nanyang Technological University(南洋理工大学)
;
The University of Sydney(悉尼大学)
;
Griffith University(格里菲斯大学)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
;
Sea AI Lab(Sea AI实验室)
;
Tsinghua University(清华大学)
;
Virginia Tech(弗吉尼亚理工大学)
;
CISPA Helmholtz Center for Information Security(CISPA海德堡信息安全部)
;
University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校)
;
Purdue University(普渡大学)
;
Duke University(杜克大学)
;
University of Wisconsin - Madison(威斯康星大学麦迪逊分校)
;
RIKEN(理化学研究所)
;
The University of Tokyo(东京大学)
机构
*
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
University of Minnesota(明尼苏达大学)
;
University of Southern California(南加州大学)
;
McGill University(麦吉尔大学)
;
Mila(Mila研究院)
;
MBZUAI(MBZUAI研究院)
Comments43 pages, 3 synthetic CV PDF's, 6 chat history PDF's and system prompts. This work was developed as part of the Responsible AI course within the Mannheim Master in Data Science (MMDS) program at the University of Mannheim
机构
*
Tribhuvan University(特里布文大学)
;
University of North Dakota(北达科他大学)
;
Youngstown State University(亚当斯州立大学)
;
University of Missouri(密苏里大学)
;
University of Toledo(托莱多大学)
CommentsAccepted to the AI-SS 2026 Workshop at the 21st European Dependable Computing Conference (EDCC 2026). To be published in the EDCC Companion Proceedings (EDCC-C)
Manipulating Multimodal Agents via Cross-Modal Prompt Injection
Le Wang, Zonghao Ying, Tianyuan Zhang, Siyuan Liang, Shengshan Hu, Mingchuan Zhang, Aishan Liu, Xianglong Liu
机构
*
Beihang University(北洋大学)
;
National University of Singapore(新加坡国立大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
Henan University of Science and Technology(河南科技大学)
机构
*
University of South Florida(佛罗里达南大学)
;
Missouri University of Science and Technology(密苏里科技大学)
;
University of Alabama(阿拉巴马大学)
;
Florida International University(佛罗里达国际大学)
;
University of Cincinnati(辛辛那提大学)
;
George Mason University(乔治·梅森大学)