BRIDGE: Benchmarking Large Language Models for Understanding Real-world Clinical Practice Text
BRIDGE:用于评估大语言模型理解现实世界临床实践文本的基准测试
Jiageng Wu, Bowen Gu, Ren Zhou, Kevin Xie, Doug Snyder, Yixing Jiang, Valentina Carducci, Richard Wyss, Rishi J Desai, Emily Alsentzer, Leo Anthony Celi, Adam Rodman, Sebastian Schneeweiss, Jonathan H. Chen, Santiago Romero-Brufau, Kueiyu Joshua Lin, Jie Yang
机构
*
Brigham and Women's Hospital(布莱根妇女医院)
;
Harvard Medical School(哈佛医学院)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Mayo Clinic(梅奥诊所)
;
Harvard T.H. Chan School of Public Health(哈佛大学陈曾熙公共卫生学院)
;
Harvard University(哈佛大学)
;
Stanford University(斯坦福大学)
;
Beth Israel Deaconess Medical Center(贝斯以色列女执事医疗中心)
;
Kempner Institute for the Study of Natural and Artificial Intelligence(肯普纳自然与人工智能研究所)
;
Broad Institute of MIT and Harvard(博德研究所)
;
Harvard Data Science Initiative(哈佛数据科学计划)
CommentsVersion 3 is focused exclusively on the first part of v1 and v2, correcting minor mathematical errors. The original co-authors have transitioned in separate follow-up works