机构
*
Department of Physics, University of Massachusetts, Amherst, USA(马萨诸塞大学物理系)
;
Department of Mathematics and Statistics, University of Massachusetts, Amherst, USA(马萨诸塞大学数学与统计学系)
;
Humboldt University of Berlin , Berlin, Germany(柏林洪堡大学)
;
Zuse Institute Berlin, Berlin, Germany(柏林祖布研究所)
From Scores to Steps: Diagnosing and Improving LLM Performance in Evidence-Based Medical Calculations
从分数到步骤:诊断和改进证据医学计算中LLM的性能
Benlu Wang, Iris Xia, Yifan Zhang, Junda Wang, Feiyun Ouyang, Shuo Han, Arman Cohan, Hong Yu, Zonghai Yao
机构
*
Department of Computer Science, Yale University, CT, USA(耶鲁大学计算机科学系)
;
Center for Healthcare Organization and Implementation Research, VA Bedford Health Care(VA贝福德医疗中心健康组织与实施研究中心)
;
Miner School of Computer and Information Sciences, UMass Lowell, MA, USA(UMass洛厄尔矿尔计算机与信息科学学院)
;
Manning College of Information and Computer Sciences, UMass Amherst, MA, USA(UMass阿默斯特马宁信息与计算机科学学院)
CommentsEqual contribution for the first two authors. To appear as an Oral presentation in the proceedings of the Main Conference on Empirical Methods in Natural Language Processing (EMNLP) 2025
机构
*
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Amazon(亚马逊)
;
University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
University at Buffalo(布法罗大学)
;
Northeastern University(东北大学)
Shape of Thought: When Distribution Matters More than Correctness in Reasoning Tasks
思维的形状:当分布比正确性更重要时在推理任务中的表现
Abhranil Chandra, Ayush Agrawal, Arian Hosseini, Sebastian Fischmeister, Rishabh Agarwal, Navin Goyal, Aaron Courville
机构
*
University of Waterloo(滑铁卢大学)
;
University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)
;
MILA - Quebec AI Institute(魁北克人工智能研究所)
;
Université de Montréal(蒙特利尔大学)
;
Microsoft Research India(微软印度研究院)
;
Google DeepMind(谷歌DeepMind)
;
Periodic Labs(周期实验室)
Knowing When to Abstain: Medical LLMs Under Clinical Uncertainty
知何时退避:医疗大语言模型在临床不确定性中的表现
Sravanthi Machcha, Sushrita Yerra, Sahil Gupta, Aishwarya Sahoo, Sharmin Sultana, Hong Yu, Zonghai Yao
机构
*
Manning College of Information and Computer Sciences, UMass Amherst, MA, USA(马萨诸塞大学阿姆赫斯特曼宁信息与计算机科学学院)
;
Center for Healthcare Organization and Implementation Research, VA Bedford Health Care(医疗组织与实施研究中心)
;
Miner School of Computer and Information Sciences, UMass Lowell, MA, USA(米纳尔计算机与信息科学学院)
CommentsEqual contribution for the first two authors; To appear in proceedings of the Main Conference of the European Chapter of the Association for Computational Linguistics (EACL) 2026
Towards AI Transparency and Accountability: A Global Framework for Exchanging Information on AI Systems
迈向人工智能透明与问责:一个全球框架用于交换人工智能系统信息
Warren Buckley, Adrian Byrne, Nicholas Perello, Cyrus Cousins, Taha Yasseri, Yair Zick, Przemyslaw Grabowicz
机构
*
University College Dublin(都柏林大学)
;
University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)
;
Duke University(杜克大学)
;
Trinity College Dublin, Technological University Dublin(都柏林圣三一学院、技术大学都柏林)
机构
*
University of Massachusetts, Amherst(马萨诸塞大学阿默斯特分校)
;
Emory University(埃默里大学)
;
University of Minnesota(明尼苏达大学)
;
University of Massachusetts, Lowell(马萨诸塞大学洛厄尔分校)
;
UMass Chan Medical School(UMass Chan医学学院)