Comments31 pages, 37 figures; submitted to Cold Regions Science and Technology on March 22, 2025; received first revision request on March 18, 2026; submitted first revision on April 7, 2026
Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning
跨商业学科的前沿人工智能性能:基于案例的知识工作和分析推理基准
Ajay Patel, Kartik Hosanagar, Ramayya Krishnan, Chris Callison-Burch, Karim Lakhani
机构
*
The Wharton School, University of Pennsylvania(宾夕法尼亚大学沃顿商学院)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Harvard Business School, Harvard University(哈佛大学哈佛商学院)
Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robustness
Infra-Bayesian 强化学习智能体在最坏情况鲁棒性上优于经典强化学习
Manish Aryal, Faiyaz Azam, Agnivo Banerjee, Syed Mahir Ahamed, Sai Sidhanth Manoharan Jayanthi, Allegra Laro, Clément Legentilhomme, Andrew Lin, Florian Lorkowski, Marina Pérez del Valle, Radman Rakhshandehroo, Patric Rommel, Emanuel Ruzak, Nathan Theng, Paul Yushin Rapoport
机构
*
Purdue University(普渡大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
WorldQuant University(WorldQuant大学)
;
UC Berkeley(加州大学伯克利分校)
;
Aix-Marseille University(阿维尼翁-马赛大学)
;
MIT(麻省理工学院)
;
University of Zurich(苏黎世大学)
;
University of British Columbia(不列颠哥伦比亚大学)
;
University of Stuttgart(斯图加特大学)
;
University of Buenos Aires(布宜诺斯艾利斯大学)
;
California State University, Fresno(弗雷斯诺加州州立大学)
;
University of Chicago(芝加哥大学)
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes
LLM推理的周期表:推理范式、方法与失败模式的结构化综述
Avinash Anand, Mahisha Ramesh, Avni Mittal, Ashutosh Kumar, Rishitej Reddy Vyalla, Erik Cambria, Zhengkui Wang, Timothy Liu, Aik Beng Ng, Simon See, Rajiv Ratn Shah
机构
*
Singapore Institute of Technology(新加坡理工大学)
;
Nvidia AI Center (SNAIC)(英伟达人工智能中心(SNAIC))
;
MIDAS Lab, IIIT Delhi(IIIT德里MIDAS实验室)
;
MIDAS Lab, IIT Mandi(IIT曼迪MIDAS实验室)
;
Owl Autonomous Imaging, Inc.(Owl自主成像公司)
;
College of Computing & Data Science, NTU Singapore(新加坡南洋理工大学计算与数据科学学院)
;
NVIDIA AI Technology Centre, Singapore(英伟达新加坡人工智能技术中心)
;
Department of Computer Science and Engineering, IIT Kanpur(IIT坎普尔计算机科学与工程系)
Comments14 pages, 1 figure, 8 tables. Corrected the DeepSeek decoding and evaluation results and updated the corresponding analysis; corrected the mathematical statement of the frozen-parameter guarantee; clarified paired statistical inference; expanded related work and corrected references
IO Factory: Simulating AI-Enabled Influence Campaigns at Scale
IO Factory:大规模模拟AI驱动的影响力行动
Lukasz Olejnik, Wenchao Dong, Jonas R. Kunst, Signe Riemer-Sørensen, Tobias Herb, Meeyoung Cha, Daniel Thilo Schroeder
机构
*
King’s College London(伦敦国王学院)
;
Max Planck Institute for Security and Privacy(马克斯·普朗克安全与隐私研究所)
;
BI Norwegian Business School(挪威商学院)
;
University of Oslo(奥斯陆大学)
;
SINTEF Digital(SINTEF数字研究所)
;
Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院)