Jailbreaking and Mitigation of Vulnerabilities in Large Language Models
大语言模型的越狱与漏洞缓解
Benji Peng, Hanxuan Chen, Keyu Chen, Qian Niu, Ziqian Bi, Ming Liu, Pohsun Feng, Tianyang Wang, Lawrence K. Q. Yan, Yizhu Wen, Yichao Zhang, Caitlyn Heqi Yin, Xinyuan Song, Riyang Bao, Jiacheng Shi
机构
*
Hunan University Changsha, PRC
;
Georgia Institute of Technology Atlanta, USA
;
Kyoto University Kyoto, Japan
;
Purdue University West Lafayette, USA
;
National Taiwan Normal University Taipei, ROC
;
University of Liverpool Suzhou, PRC
;
Hong Kong University of Science
;
University of Hawaii Honolulu, USA
;
The University of Texas at Dallas Dallas, USA
;
University of Wisconsin-Madison Madison, USA
;
Emory University Atlanta, USA
;
College of William \& Mary Williamsburg, USA
机构
*
John A. Paulson School of Engineering And Applied Sciences, Harvard University(哈佛大学约翰·A·保罗森工程与应用科学学院)
;
Department of Brain and Cognitive Sciences, Massachusetts Institute of Technology(麻省理工学院脑科学与认知科学系)
;
Speech and Hearing Bioscience and Technology, Harvard Medical School(哈佛医学院语音与听力生物科学与技术系)
;
Kempner Institute for the Study of Natural and Artificial Intelligence, Harvard University(哈佛大学自然与人工智能研究学院)
;
Center for Brain Science, Harvard University(哈佛大学脑科学中心)
Agora: Toward Autonomous Bug Detection in Production-Level Consensus Protocols with LLM Agents
Agora: 面向生产级共识协议中自主漏洞检测的LLM智能体
Xiang Liu, Sa Song, Zhaowei Zhang, Huiying Lan, Jason Zeng, Ming Wu, Michael Heinrich, Yong Sun, Ceyao Zhang
机构
*
School of Computing, National University of Singapore(新加坡国立大学计算机学院)
;
School of Information and Telecommunication Engineering, Beijing University of Posts and Telecommunications(北京邮电大学信息与电信工程学院)
;
Peking University(北京大学)
;
G Labs(0G实验室)