A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework
对LLM-as-a-Judge在医疗领域的综述及MedJUDGE框架
Chenyu Li, Zohaib Akhtar, Mingu Kwak, Yuelyu Ji, Hang Zhang, Tracey Obi, Yufan Ren, Xizhi Wu, Sonish Sivarajkumar, Harold P. Lehmann, Shyam Visweswaran, Michael J. Becich, Danielle L. Mowery, Renxuan Liu, Haoyang Sun, Yanshan Wang
机构
*
Department of Biomedical Informatics, School of Medicine, University of Pittsburgh(匹兹堡大学医学院生物医学信息学系)
;
Department of Health Information Management, School of Health and Rehabilitation Sciences, University of Pittsburgh(匹兹堡大学健康与康复科学学院健康信息管理系)
;
OpenCura, Health Innovation Consortium(OpenCura健康创新联盟)
;
Northwestern University, Kellogg School of Management(西北大学凯洛格管理学院)
;
Intelligent Systems Program, School of Computing and Information, University of Pittsburgh(匹兹堡大学计算与信息学院智能系统项目)
;
Johns Hopkins University School of Medicine Biomedical Informatics and Data Science(约翰霍普金斯大学医学院生物医学信息学与数据科学)
;
Clinical and Translational Science Institute, University of Pittsburgh(匹兹堡大学临床与转化科学研究所)
;
Institute for Biomedical Informatics, University of Pennsylvania(宾夕法尼亚大学生物医学信息学研究所)
;
Data Science, School of Computing and Information, University of Pittsburgh(匹兹堡大学计算与信息学院数据科学)
The Consensus Trap: Dissecting Subjectivity and the "Ground Truth" Illusion in Data Annotation
共识陷阱:数据标注中主观性与‘真实真相’幻觉的剖析
Sheza Munir, Benjamin Mah, Krisha Kalsi, Shivani Kapania, Julian Posada, Edith Law, Ding Wang, Syed Ishtiaque Ahmed
机构
*
University of Toronto Computer Science(多伦多大学计算机科学系)
;
University of Toronto Engineering Science(多伦多大学工程科学系)
;
Carnegie Mellon University School of Computer Science(卡内基梅隆大学计算机科学学院)
;
Yale University American Studies(耶鲁大学美国研究系)
;
University of Waterloo Computer Science(滑铁卢大学计算机科学系)
;
Google Research(谷歌研究)
;
University of Toronto(多伦多大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Yale University(耶鲁大学)
;
University of Waterloo(滑铁卢大学)
AEGIS: An Operational Infrastructure for Post-Market Governance of Adaptive Medical AI Under US and EU Regulations
AEGIS:一种用于美国和欧盟法规下适应性医疗AI市场后治理的操作基础设施
Fardin Afdideh, Mehdi Astaraki, Fernando Seoane, Farhad Abtahi
机构
*
Department of Clinical Science, Intervention and Technology, Karolinska Institutet(临床科学、干预与技术部门,Karolinska研究院)
;
Department of Medical Radiation Physics, Stockholm University(医学辐射物理学部门,斯德哥尔摩大学)
;
Department of Oncology-Pathology, Karolinska Institutet(肿瘤学-病理学部门,Karolinska研究院)
;
Department of Clinical Physiology, Karolinska University Hospital(临床生理学部门,Karolinska大学医院)
;
Department of Textile Technology, University of Bor s(纺织技术部门,Bor s大学)
;
Department of Medical Technologies, Karolinska University Hospital(医学技术部门,Karolinska大学医院)
;
Department of Biomedical Engineering and Health System, KTH Royal Institute of Technology(生物医学工程与健康系统部门,KTH皇家理工学院)
Credibility Governance: A Social Mechanism for Collective Self-Correction under Weak Truth Signals
可信治理:在弱真相信号下的一种社会机制,用于集体自我校正
Wanying He, Yanxi Lin, Ziheng Zhou, Xue Feng, Min Peng, Qianqian Xie, Zilong Zheng, Yipeng Kang
机构
*
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
;
Tsinghua University(清华大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,BIGAI)
Buy versus Build an LLM: A Decision Framework for Governments
买还是建一个大语言模型:政府的决策框架
Jiahao Lu, Ziwei Xu, William Tjhi, Junnan Li, Antoine Bosselut, Pang Wei Koh, Mohan Kankanhalli
机构
*
National University of Singapore(新加坡国立大学)
;
AI Singapore(AI新加坡)
;
Salesforce AI Research(Salesforce AI研究)
;
EPFL(苏黎世联邦理工学院)
;
University of Washington(华盛顿大学)
;
Allen Institute for AI(人工智能研究院)
Computational Basis of LLM's Decision Making in Social Simulation
大语言模型在社会模拟中的决策机制计算基础
Ji Ma
机构
*
LBJ School of Public Affairs, University of Texas at Austin(德克萨斯大学奥斯汀分校公共事务学院LBJ学院)
;
Gradel Institute of Charity, New College, University of Oxford(牛津大学格拉德尔慈善研究所)
From Text to Multimodality: Exploring the Evolution and Impact of Large Language Models in Medical Practice
从文本到多模态:探索大型语言模型在医疗实践中的演变与影响
Qian Niu, Keyu Chen, Ming Li, Pohsun Feng, Ziqian Bi, Lawrence KQ Yan, Yichao Zhang, Caitlyn Heqi Yin, Cheng Fei, Junyu Liu, Tianyang Wang, Yunze Wang, Silin Chen, Ming Liu, Benji Peng, Xinyuan Song, Ziyuan Qin, Riyang Bao, Zekun Jiang
机构
*
Kyoto University(京都大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
National Taiwan Normal University(台湾师范大学)
;
Indiana University(印第安纳大学)
;
Hong Kong University of Science(香港科学大学)
;
The University of Texas at Dallas(德克萨斯大学达拉斯分校)
;
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
;
Cornell University(康奈尔大学)
;
University of Liverpool(利物浦大学)
;
University of Edinburgh(爱丁堡大学)
;
Zhejiang University(浙江大学)
;
Purdue University(Purdue 大学)
;
Emory University, Atlanta, GA, USA(埃默里大学)
;
West China Biomedical Big Data Center, West China Hospital, Sichuan University, Chengdu, China(西京生物大数据中心,四川大学西京医院,成都,中国)
Alif: Advancing Urdu Large Language Models via Multilingual Synthetic Data Distillation
Muhammad Ali Shafique, Kanwal Mehreen, Muhammad Arham, Maaz Amjad, Sabur Butt, Hamza Farooq
机构
*
University of British Columbia(不列颠哥伦比亚大学)
;
Texas Tech University(德克萨斯技术大学)
;
Institute for the Future of Education, Tecnológico de Monterrey(教育未来研究所,墨西哥蒙特雷技术学院)