机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
Zhejiang University(浙江大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究所)
;
Nanjing University(南京大学)
;
Shanghai Innovation Institute(上海创新研究院)
CommentsFull version of a to be published paper in proceedings of the 9th AAAI/ACM conference in AI, Ethics and Society (AIES 2026). Includes supplementary material. 20 pages, 2 figures
Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
真实还是编造?利用因果归因减轻解释中的奖励作弊
Pedro Ferreira, Wilker Aziz, Ivan Titov
机构
*
Institute for Logic, Language and Computation (ILLC), University of Amsterdam(逻辑、语言与计算研究所(ILLC),阿姆斯特丹大学)
;
Institute for Language, Cognition and Computation (ILCC), University of Edinburgh(语言、认知与计算研究所(ILCC),爱丁堡大学)
LearNAT: Learning NL2SQL with AST-guided Task Decomposition for Large Language Models
LearNAT: 基于AST引导任务分解的NL2SQL大语言模型学习
Weibin Liao, Xin Gao, Tianyu Jia, Rihong Qiu, Yifan Zhu, Yang Lin, Xinyu Ma, Junfeng Zhao, Yasha Wang
机构
*
School of Computer Science, Peking University(北京大学计算机科学系)
;
Key Laboratory of High Confidence Software Technologies, Ministry of Education(教育部高可信软件技术重点实验室)
;
Big Data Technology Research Center, Nanhu Laboratory(纳米实验室大数据技术研究中心)
;
National Engineering Research Center For Software Engineering, Peking University(北京大学软件工程国家工程研究中心)
;
Peking University Information Technology Institute (Tianjin Binhai)(北京大学信息技术研究院(天津滨海))
;
School of Computer Sciences, Beijing University of Posts and Telecommunications(北京邮电大学计算机科学系)
;
Huawei Technologies Co., Ltd(华为技术有限公司)
;
Seed, ByteDance Inc.(字节跳动公司)
机构
*
Center for Advanced Intelligence Project, RIKEN(日本理化学研究所先进智能研究中心)
;
Graduate School of Frontier Sciences, The University of Tokyo(东京大学前沿科学研究生院)
;
Northeastern University(东北大学)
;
Zhejiang Gongshang University(浙江工商大学)
Patches of Nonlinearity: Instruction Vectors in Large Language Models
非线性补丁:大型语言模型中的指令向量
Irina Bigoulaeva, Jonas Rohweder, Subhabrata Dutta, Iryna Gurevych
机构
*
Ubiquitous Knowledge Processing Lab (UKP Lab)(通用知识处理实验室)
;
Department of Computer Science, Technical University of Darmstadt and National Research Center for Applied Cybersecurity ATHENE, Germany(计算机科学系,达姆施塔特技术大学和应用网络安全国家研究中心ATHENE,德国)
BayLing-Duplex: Native Full-Duplex Speech Dialogue with a Single Autoregressive LLM
BayLing-Duplex: 单一自回归LLM的原生全双工语音对话
Qingkai Fang, Shoutao Guo, Yang Feng
机构
*
Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, Chinese Academy of Sciences (ICT/CAS)(中国科学院计算技术研究所智能信息处理重点实验室)
;
Key Laboratory of AI Safety, Chinese Academy of Sciences(中国科学院人工智能安全重点实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
Andy Wang, Parv Mahajan, David Demitri Africa, Alexandra Souly, Jordan Taylor, Robert Kirk
机构
*
Constellation University of Wisconsin-Madison(威斯康星大学麦迪逊分校星座研究所)
;
Constellation Georgia Institute of Technology(佐治亚理工学院星座研究所)
;
UK AI Security Institute(英国人工智能安全研究所)