PrisonBreak: Jailbreaking Large Language Models with at Most Twenty-Five Targeted Bit-flips
Zachary Coalson, Jeonghyun Woo, Chris S. Lin, Joyce Qu, Yu Sun, Shiyang Chen, Lishan Yang, Gururaj Saileshwar, Prashant Nair, Bo Fang, Sanghyun Hong
机构
*
Oregon State University(俄勒冈州立大学)
;
University of British Columbia(不列颠哥伦比亚大学)
;
University of Toronto(多伦多大学)
;
George Mason University(乔治·梅森大学)
;
Rutgers University(罗格斯大学)
;
University of Texas at Arlington(德克萨斯大学阿灵顿分校)
机构
*
School of Electronic, Electrical and communication Engineering, UCAS, Beijing.(电子电气与通信工程学院,中国科学院大学,北京)
;
Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, CAS, Beijing.(智能信息处理重点实验室,计算技术研究所,中国科学院,北京)
;
School of Computer Science and Technology, UCAS, Beijing.(计算机科学与技术学院,中国科学院大学,北京)
;
Key Laboratory of Big Data Mining and Knowledge management, UCAS, Beijing(大数据挖掘与知识管理重点实验室,中国科学院大学,北京)
;
Nanyang Technological University, Singapore.(南洋理工大学,新加坡)
;
Shenzhen Campus of Sun Yat-sen University, Shenzhen.(孙中山大学深圳校区,深圳)
机构
*
Tsinghua Shenzhen International Graduate School(清华大学深圳国际研究生院)
;
Pengcheng Laboratory(鹏城实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
Jilin University(吉林大学)
;
Southwest University(西南大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Shenzhen University(深圳大学)