机构
*
Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所)
;
Science & Technology on Integrated Information System Laboratory(信息集成信息系统实验室)
;
State Key Laboratory of Complex System Modeling and Simulation Technology(复杂系统建模与仿真技术国家重点实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Singapore Management University(新加坡管理大学)
Why Does Reinforcement Learning Generalize? A Feature-Level Mechanistic Study of Post-Training in Large Language Models
为什么强化学习能够泛化?对大语言模型后训练的特征层面机制研究
Dan Shi, Zhuowen Han, Simon Ostermann, Renren Jin, Josef van Genabith, Deyi Xiong
机构
*
TJUNLP Lab, School of Computer Science and Technology, Tianjin University(天津大学计算机科学与技术学院TJUNLP实验室)
;
German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)
;
Saarland University(萨尔兰大学)
机构
*
Interdisciplinary Program in Artificial Intelligence, Seoul National University(人工智能交叉学科项目,首尔国立大学)
;
Electronics and Telecommunications Research Institute(电子与电信研究所)
Mixture of Heterogeneous Grouped Experts for Language Modeling
异质分组专家的混合模型用于语言建模
Zhicheng Ma, Xiang Liu, Zhaoxiang Liu, Ning Wang, Yi Shen, Kai Wang, Shuming Shi, Shiguo Lian
机构
*
Data Science & Artificial Intelligence Research Institute, China Unicom(中国unicom数据科学与人工智能研究院)
;
Unicom Data Intelligence, China Unicom(中国unicom数据智能)
Comments15 pages, 7 figures; minor source and formatting cleanup; results unchanged
Journal refProceedings of the Ninth Fact Extraction and VERification Workshop (FEVER), pages 59-73, Rabat, Morocco. Association for Computational Linguistics, 2026
Hiba Arnaout, Anmol Goel, H. Andrew Schwartz, Steffen T. Eberhardt, Dana Atzil-Slonim, Gavin Doherty, Brian Schwartz, Wolfgang Lutz, Tim Althoff, Munmun De Choudhury, Hamidreza Jamalabadi, Raj Sanjay Shah, Flor Miriam Plaza-del-Arco, Dirk Hovy, Maria Liakata, Iryna Gurevych
机构
*
Technische Universität Darmstadt(德累斯顿技术大学)
;
Vanderbilt University(范德比尔特大学)
;
Trier University(特里尔大学)
;
Bar-Ilan University(巴伊兰大学)
;
Trinity College Dublin(都柏林三一学院)
;
University of Washington(华盛顿大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Phillips-Universität Marburg(马尔堡菲利普大学)
;
LIACS, Leiden University(莱顿大学LIACS研究中心)
;
Bocconi University(博科尼大学)
;
Queen Mary University London(伦敦玛丽女王大学)
;
Alan Turing Institute(艾伦·图灵研究所)
Journal refProceedings of the 19th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers), pages 8233-8264, 2026
机构
*
Department of Statistics, LMU Munich(慕尼黑大学统计系)
;
Heidelberg Academy of Sciences and Humanities(海德堡科学院和人文科学学院)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心(MCML))
OMHBench: Benchmarking Balanced and Grounded Omni-Modal Multi-Hop Reasoning
OMHBench: 多模态多跳推理的基准测试
Seunghee Kim, Ingyu Bang, Seokgyu Jang, Changhyeon Kim, Sanghwan Bae, Jihun Choi, Richeng Xuan, Taeuk Kim
机构
*
Hanyang University(翰阳大学)
;
NAVER Cloud(NAVER云)
;
Knowledge Work Inc.(知识工作公司)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
Sony AI(索尼人工智能)
Large Language Models Are Effective Human Annotation Assistants, But Not Good Independent Annotators
大语言模型是有效的标注助手,但不是好的独立标注者
Feng Gu, Zongxia Li, Carlos Rafael Colon, Benjamin Evans, Ishani Mondal, Jordan Lee Boyd-Graber
机构
*
Department of Computer Science, University of Maryland(大学计算机科学系)
;
National Consortium for the Study of Terrorism and Responses to Terrorism(反恐与反恐响应国家研究中心)