Reasoning Errors Have a Region and a Direction in the Residual-Stream Trajectory of LLMs
推理错误在大语言模型(LLMs)残差流轨迹中具有特定区域与方向
Hamed Damirchi, Ignacio Meza De la Jara, Damith Ranasinghe, Yuhang Liu, Javen Shi
机构
*
Australian Institute for Machine Learning(澳大利亚机器学习研究所)
;
Adelaide University(阿德莱德大学)
;
Naval Group Pacific(太平洋海军集团)
;
Responsible AI Research Centre(负责任人工智能研究中心)
SP-Mind: An Autonomous Reasoning Agent for Spatial Proteomics Analysis
SP-Mind: 用于空间蛋白质组学分析的自主推理智能体
Yucheng Yuan, Yuanfeng Ji, Zhongxiao Li, Ruijiang Li
机构
*
Department of Computer Science, Stanford University, Stanford, USA(计算机科学系,斯坦福大学,斯坦福,美国)
;
Department of Radiation Oncology, Stanford University, Stanford, USA(放射肿瘤学系,斯坦福大学,斯坦福,美国)
BenHalluEval: A Multi-Task Hallucination Evaluation Framework for Large Language Models on Bengali
BenHalluEval:孟加拉语大语言模型的多任务幻觉评估框架
Shefayat E Shams Adib, Ahmed Alfey Sani, Ekramul Alam Esham, Ajwad Abrar, Ishmam Tashdeed, Md Taukir Azam Chowdhury
机构
*
Department of Computer Science and Engineering, Islamic University of Technology(伊斯兰科技大学计算机科学与工程系)
;
Department of Computer Science and Engineering, University of California(加州大学计算机科学与工程系)
机构
*
School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院)
;
NiuTrans Research(NiuTrans研究院)
;
Institute of Psychology, CAS(中国科学院心理研究所)
;
Kunming University of Science and Technology(昆明理工大学)
机构
*
The University of Hong Kong(香港大学)
;
Nanjing University(南京大学)
;
University of Science and Technology of China(中国科学技术大学)
;
National University of Singapore(新加坡国立大学)
;
Fudan University(复旦大学)
Comments6 figures and 5 tables. Hong Jiang, Junnan Zhu, and Jingwang Huang contributed equally. Jiang Zhong and Kaiwen Wei are corresponding authors. Code and data are available at https://github.com/hongshi4/M3R-Bench