Thought Purity: A Defense Framework For Chain-of-Thought Attack
思维纯洁性:针对思维链攻击的防御框架
Zihao Xue, Zhen Bi, Long Ma, Zhenlin Hu, Yan Wang, Xueshu Chen, Zhenfang Liu, Kang Zhao, Jie Xiao, Jungang Lou
机构
*
Huzhou University(湖州大学)
;
Zhejiang Key Laboratory of Intelligent Education Technology and Application(浙江智能教育技术与应用重点实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
Alibaba Group(阿里巴巴集团)
;
Zhejiang University of Technology(浙江工业大学)
Reasoning Large Language Model Errors Arise from Hallucinating Critical Problem Features
Alex Heyman, Joel Zylberberg
机构
*
Department of Electrical Engineering and Computer Science(电气工程与计算机科学系)
;
York University(约克大学)
;
Jules Stein Eye Institute(朱利斯·斯坦眼科研究所)
;
University of California(加州大学)
;
University of California Los Angeles(加州大学洛杉矶分校)
Rethinking the Chain-of-Thought: The Roles of In-Context Learning and Pre-trained Priors
Hao Yang, Zhiyu Yang, Yunjie Zhang, Shanyi Zhu, Lin Yang
机构
*
School of Intelligence Science and Technology, National Key Laboratory for Novel Software Technology, Nanjing University(智能科学与技术学院,新型软件技术国家重点实验室,南京大学)
;
School of Computing and Information Systems, Singapore Management University(计算与信息系统学院,新加坡管理大学)
;
Central South University(中南大学)
;
School of Global Education and Development, International Chinese Language Education, University of Chinese Academy of Social Sciences(全球教育与发展学院,国际中文教育,中国社会科学院)
CommentsAccepted for The 2024 Conference on Empirical Methods in Natural Language Processing (EMNLP 2024). Final version includes additional models and additional inference patterns