Enhancing LLMs' Clinical Reasoning with Real-World Data from a Nationwide Sepsis Registry
利用全国脓毒症登记处的真实世界数据增强大语言模型的临床推理能力
Junu Kim, Chaeeun Shim, Sungjin Park, Su Yeon Lee, Gee Young Suh, Chae-Man Lim, Seong Jin Choi, Song Mi Moon, Kyoung-Ho Song, Eu Suk Kim, Hong Bin Kim, Sejoong Kim, Chami Im, Dong-Wan Kang, Yong Soo Kim, Hee-Joon Bae, Sung Yoon Lim, Han-Gil Jeong, Edward Choi
机构
*
Korea Advanced Institute of Science and Technology(韩国科学技术院)
;
Microsoft(微软)
;
Asan Medical Center, University of Ulsan College of Medicine(釜山大学医学院阿桑医疗中心)
;
Samsung Medical Center, Sungkyunkwan University School of Medicine(成均馆大学医学院三星医疗中心)
;
Seoul National University Bundang Hospital, Seoul National University College of Medicine(首尔国立大学医学院首尔国立大学医院)
Evaluating Retrieval-Augmented Generation vs. Long-Context Input for Clinical Reasoning over EHRs
评估检索增强生成与长上下文输入用于电子健康记录临床推理的效果
Skatje Myers, Dmitriy Dligach, Timothy A. Miller, Samantha Barr, James Landefeld, Yanjun Gao, Matthew Churpek, Anoop Mayampurath, Majid Afshar
机构
*
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
;
Loyola University Chicago(芝加哥洛约拉大学)
;
Boston Children’s Hospital Harvard Medical School(波士顿儿童医院哈佛医学院)
;
University of Colorado-Anschutz(科罗拉多大学安舒茨分校)
Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework
衡量LLM中的推理质量:一个多维行为框架
Ali Şenol, Garima Agrawal, Huan Liu
机构
*
Department of Computer Engineering, Tarsus University(塔鲁斯大学计算机工程系)
;
School of Computing and Augmented Intelligence (SCAI), Arizona State University (ASU)(计算与增强智能学院(SCAI),亚利桑那州立大学(ASU))
;
HumaConn AI Consulting(HumaConn AI咨询)
BenGER: Benchmarking LLM Systems on Subsumption-Based Legal Reasoning in German Law
BenGER:德国法律中基于归入的法律推理的LLM系统基准测试
Sebastian Nagl, Ann-Kristin Mayrhofer, Martin Heidebach, Aleyna Koçak, Anne Zettelmeier, Elly Breu, Angelina Greiner, Sofija Milijas, Matthias Grabmair
机构
*
Technical University of Munich (TUM)(慕尼黑技术大学)
;
Ludwig Maximilian University of Munich (LMU)(慕尼黑路德维希-马克西米利安大学)
;
University of Konstanz(康斯坦茨大学)
;
University of Saarbrücken(萨尔布吕肯大学)
CycliST: A Video Language Model Benchmark for Reasoning on Cyclical State Transitions
CycliST:用于循环状态转换推理的视频语言模型基准
Simon Kohaut, Daniel Ochs, Shun Zhang, Benedict Flade, Julian Eggert, Kristian Kersting, Devendra Singh Dhami
机构
*
Artificial Intelligence and Machine Learning Lab, TU Darmstadt(人工智能与机器学习实验室,图腾斯达特技术大学)
;
Konrad Zuse School of Excellence in Learning and Intelligent Systems (ELIZA)(Konrad Zuse 学校(ELIZA))
;
Honda Research Institute Europe GmbH, Offenbach, Germany(本田欧洲研究院,奥芬巴赫,德国)
;
Uncertainty in Artificial Intelligence Group, TU Eindhoven(人工智能不确定性小组,埃因霍温技术大学)
;
Hessian Center for AI (hessian.AI)(黑森人工智能中心(hessian.AI))
;
Center for Cognitive Science(认知科学中心)
;
German Center for Artificial Intelligence (DFKI)(德国人工智能中心(DFKI))
Can LLMs Accurately Score Medical Diagnoses and Clinical Reasoning?
LLM能否准确评分医学诊断和临床推理?
Amy Rouillard, Sitwala Mundia, Linda Camara, Ziyaad Dangor, Michael Cameron Gramanie, Ismail Kalla, Shabir A. Madhi, Kajal Morar, Marlvin T. Ncube, Haroon Saloojee, Bruce A. Bassett
机构
*
Wits MIND Institute, University of the Witwatersrand, Johannesburg, South Africa(维特士心理研究所,沃斯兰德大学,约翰内斯堡,南非)
;
Grai Labs, Cape Town, South Africa(格雷实验室,开普敦,南非)
;
South African Medical Research Council Vaccines and Infectious Diseases Analytics Research Unit, Faculty of Health Sciences, University of the Witwatersrand, Johannesburg, South Africa(南非医学研究理事会疫苗和传染病分析研究组,健康科学学院,沃斯兰德大学,约翰内斯堡,南非)
;
Department of Internal Medicine, Charlotte Maxeke Johannesburg Academic Hospital, and Faculty of Health Sciences, University of the Witwatersrand, Johannesburg, South Africa(内科学系,查理·马克斯凯约翰内斯堡学术医院,以及健康科学学院,沃斯兰德大学,约翰内斯堡,南非)
;
Department of Paediatrics and Child Health, Faculty of Health Sciences, University of the Witwatersrand, Johannesburg, South Africa(儿科学与儿童健康系,健康科学学院,沃斯兰德大学,约翰内斯堡,南非)
;
Wits MIND Institute, University of the Witwatersrand, Johannesbu(维特士心理研究所,沃斯兰德大学,约翰内斯堡)
Comments9 pages main text, 31 pages total (including references and appendix). 5 figures, 16 tables. Preprint under review. Code and data will be made available upon publication
机构
*
Department of Computer Science, National Centre for Text Mining, The University of Manchester(计算机科学系,国家文本挖掘中心,曼彻斯特大学)
;
ELLIS Manchester(曼彻斯特ELLIS)
;
School of Computing, Queen’s University, Ontario, Canada(计算学院,加拿大皇后大学)
;
Computer Science, University of Illinois Chicago(计算机科学,伊利诺伊大学芝加哥分校)
;
ELLIS Institute Finland(芬兰ELLIS研究所)
;
University of Turku(图尔库大学)
;
Department of Computer Science and Artificial Intelligence, Umm Al-Qura University, Makkah, Saudi Arabia(计算机科学与人工智能系,乌姆·阿勒·卡拉大学,麦加,沙特阿拉伯)
CommentsThe paper is withdrawn for further clarification of the alignment between the proposed knowledge transfer framework and its implementation, and for refinement of the transfer span definition and experimental evaluation design
Teaching agentic AI to learn expert reasoning for rare disease diagnosis
LiteOdyssey: 一种用于可解释罕见病诊断的轻量级推理AI智能体
Minh-Ha Nguyen, Erica Gray, Bryce A. Schuler, Kevin W. Byram, Chih-Ting Yang, Fan Ma, Hua Xu, Wu-Chen Su, Chao Yan, Wei-Qi Wei, Adam Wright, Lisa Bastarache, Josh Peterson, Lingyao Li, Siyuan Ma, Undiagnosed Diseases Network, Rizwan Hamid, Thomas A. Cassini, Cathy Shyr
机构
*
Vanderbilt University(范德堡大学)
;
Vanderbilt University Medical Center(范德堡大学医学中心)
;
University of South Florida(南佛罗里达大学)
Comments50 pages, 4 figures. v3: expanded to a unified 69-paper corpus through July 31, 2026; adds restored-state identification results, blind coding of a 42-paper full-text subset, a source-located reporting audit, the CA-ID Card, and replay-fidelity analysis
机构
*
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
;
Center for Language and Information Research, Wuhan University(武汉大学语言信息研究中心)
;
The University of Manchester(曼彻斯特大学)
;
School of Computer Science, Wuhan University(武汉大学计算机科学学院)
;
Mount Holyoke College(马歇尔学院)
;
Hefei University of Technology(合肥工业大学)
机构
*
Department of Data Science and Hong Kong Institute of AI for Science, City University of Hong Kong(数据科学系和香港人工智能科学研究所,香港城市大学)
;
Li Auto Inc., China(中国利汽车公司)
;
Department of Statistics, University of Oxford(统计系,牛津大学)
Ruiqi Wu, Yuang Yao, Tengfei Ma, Chenran Zhang, Na Su, Tao Zhou, Geng Chen, Wen Fan, Yi Zhou
机构
*
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
Department of Ophthalmology, The First Affiliated Hospital of Nanjing Medical University(南京医科大学第一附属医院眼科学系)
;
School of Computer Science and Engineering, Nanjing University of Science and Technology(南京理工大学计算机科学与工程学院)
;
School of Computer Science, Northwestern Polytechnical University(西北工业大学计算机学院)
RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension
RefBench-PRO:面向感知与推理的指称表达理解基准
Tianyi Gao, Hao Li, Han Fang, Xin Wei, Xiaodong Dong, Hongbo Sun, Ye Yuan, Zhongjiang He, Jinglin Xu, Jingmin Xin, Hao Sun
机构
*
National Key Laboratory of Human-Machine Hybrid Augmented Intelligence(国家人机混合增强智能重点实验室)
;
National Engineering Research Center for Visual Information and Applications(国家视觉信息与应用工程技术研究中心)
;
Institute of Artificial Intelligence and Robotics(人工智能与机器人研究院)
;
Xi’an Jiaotong University(西安交通大学)
;
Institute of Artificial Intelligence (TeleAI)(人工智能研究院(TeleAI))
;
China Telecom(中国电信)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Science and Technology Beijing(北京科技大学)
Enhancing Pathological VLMs with Cross-scale Reasoning
增强病理视觉语言模型的跨尺度推理能力
Chi Phan, Tianyi Zhang, Qiaochu Xue, Yufeng Wu, Dan Hu, Zeyu Liu, Sudong Wang, Yueming Jin
机构
*
Department of Electrical and Computer Engineering, National University of Singapore(新加坡国立大学电气与计算机工程系)
;
PuzzleLogic Pte Ltd(PuzzleLogic私人有限公司)
;
Department of Pathology, Fujian Medical University Cancer Hospital & Fujian Cancer Hospital(福建医科大学附属肿瘤医院病理科暨福建省肿瘤医院)