arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12554 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12554 篇

2402.10828 2026-03-09 cs.RO cs.AI 88%

RAG-Driver: Generalisable Driving Explanations with Retrieval-Augmented In-Context Learning in Multi-Modal Large Language Model

RAG-Driver: 多模态大语言模型中基于检索的可泛化驾驶解释

Jianhao Yuan, Shuyang Sun, Daniel Omeiza, Bo Zhao, Paul Newman, Lars Kunze, Matthew Gadd

机构 * University of Oxford(牛津大学) Beijing Academy of Artificial Intelligence(北京人工智能研究院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 RAG-Driver通过检索增强的上下文学习,实现高性能、可解释和可泛化的自动驾驶系统。

Comments 14 pages, 6 figures

Journal ref Robotics: Science and Systems (RSS) 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15339 2026-02-27 cs.CL 88%

DeVisE: Behavioral Testing of Medical Large Language Models

DeVisE:医疗大语言模型的行为测试

Camila Zurdo Tagliabue, Heloisa Oss Boll, Aykut Erdem, Erkut Erdem, Iacer Calixto

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 DeVisE通过受控反事实测试评估医疗大语言模型的行为,揭示其在临床决策中的表现差异。

Comments Camera-ready version published at Findings of the EACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16824 2026-02-24 cs.LG 88%

Predicting New Research Directions in Materials Science using Large Language Models and Concept Graphs

利用大语言模型和概念图预测材料科学的新研究方向

Thomas Marwitz, Alexander Colsmann, Ben Breitung, Christoph Brabec, Christoph Kirchlechner, Eva Blasco, Gabriel Cadilha Marques, Horst Hahn, Michael Hirtz, Pavel A. Levkin, Yolita M. Eggeler, Tobias Schlöder, Pascal Friederich

机构 * Institute of Theoretical Informatics, Karlsruhe Institute of Technology(理论信息学院,卡尔斯鲁厄技术大学) Material Research Center for Energy Systems, Karlsruhe Institute of Technology(能源系统材料研究中心,卡尔斯鲁厄技术大学) Institute of Nanotechnology, Karlsruhe Institute of Technology(纳米技术学院,卡尔斯鲁厄技术大学) Department of High Throughput Methods in Photovoltaics, Forschungszentrum Jülich GmbH(光伏高通量方法部门,焦耳研究中心 GmbH) Department of Materials Science and Engineering, Institute of Materials for Electronics and Energy Technology (i-MEET), Friedrich-Alexander-Universität Erlangen-Nürnberg(材料科学与工程系,电子与能源技术材料研究所(i-MEET),埃尔朗根-纽伦堡弗里德里希-亚历山大大学) Institute for Applied Materials, Karlsruhe Institute of Technology(应用材料研究所,卡尔斯鲁厄技术大学) Institute for Molecular Systems Engineering and Advanced Materials, Heidelberg University(分子系统工程与先进材料研究所,海德堡大学) Department of Materials Science and Engineering, University of Arizona(材料科学与工程系,亚利桑那大学) Institute of Biological and Chemical Systems – Functional Molecular Systems(生物与化学系统研究所——功能分子系统)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 本研究利用大语言模型和概念图,通过分析科学文献中的语义信息,预测材料科学领域的新研究方向,提升科研创新效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13742 2026-02-24 cs.CV cs.AI 88%

DL$^3$M: A Vision-to-Language Framework for Expert-Level Medical Reasoning through Deep Learning and Large Language Models

DL$^3$M: 一种通过深度学习和大语言模型实现专家级医学推理的视觉-语言框架

Md. Najib Hasan, Imran Ahmad, Sourav Basak Shuvo, Md. Mahadi Hasan Ankon, Sunanda Das, Nazmul Siddique, Hui Wang

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 DL$^3$M通过结合深度学习和大语言模型,实现专家级医学推理,但当前LLMs在高风险医疗决策中仍不可靠。

Comments This work was submitted without the consent of my current adviser. Additionally, it overlaps with my unpublished research work. In order to avoid potential academic and authorship conflicts, I am requesting withdrawal of the paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18880 2026-02-24 cs.CV cs.AI 88%

FOCA: Frequency-Oriented Cross-Domain Forgery Detection, Localization and Explanation via Multi-Modal Large Language Model

FOCA:基于多模态大语言模型的频率导向跨域伪造检测、定位与解释

Zhou Liu, Tonghua Su, Hongshi Zhang, Fuxiang Yang, Donglin Di, Yang Song, Lei Fan

机构 * Harbin Institute of Technology(哈尔滨工业大学) DZ-Matrix Guangdong Laboratory of Artificial Intelligence and Digital Economy(广东省人工智能与数字经济实验室) Chongqing Research Institute of HIT(哈尔滨工业大学重庆研究院) University of New South Wales(新南威尔士大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 FOCA通过多模态大语言模型整合空间与频域特征,实现图像伪造的高精度检测、定位及可解释性解释,优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13860 2026-02-17 cs.CL 88%

Tutoring Large Language Models to be Domain-adaptive, Precise, and Safe

指导大语言模型成为领域自适应、精确和安全的

Somnath Banerjee

机构 * Indian Institute of Technology Kharagpur(印度理工学院克哈格浦分校)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本研究通过领域适应、伦理严谨性和文化对齐,指导大语言模型实现精确性、安全性和全球包容性。

Comments Accepted to the PhD Symposium at Web Conference 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01159 2026-02-16 cs.CL 88%

Large Language Models for Healthcare Text Classification: A Systematic Review

用于医疗文本分类的大型语言模型:系统综述

Hajar Sakai, Sarah S. Lam

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文系统综述了LLMs在医疗文本分类中的应用,分析了现有研究的现状、挑战及未来研究方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04392 2026-02-05 cs.CL 88%

Evaluating the Presence of Sex Bias in Clinical Reasoning by Large Language Models

通过大语言模型评估临床推理中的性别偏见存在性

Isabel Tsintsiper, Sheng Wong, Beth Albert, Shaun P Brennecke, Gabriel Davis Jones

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 研究发现大语言模型在临床推理中存在性别偏见,不同模型表现不同,需谨慎配置与监督以确保医疗应用的安全性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.16419 2026-02-05 cs.CL cs.CV 88%

Learning Domain Knowledge in Multimodal Large Language Models through Reinforcement Fine-Tuning

通过强化微调学习多模态大语言模型中的领域知识

Qinglong Cao, Yuntian Chen, Chao Ma, Xiaokang Yang

机构 * MoE Key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University, Shanghai, China(人工智能大规模并行计算实验室,人工智能研究院,上海交通大学,上海)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文提出通过强化微调框架在优化层面整合领域知识,提升多模态大语言模型在专门领域任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00041 2026-02-04 cs.CY cs.AI cs.HC 88%

Student Perceptions of Large Language Models Use in Self-Reflection and Design Critique in Architecture Studio

学生对在建筑工作室中使用大型语言模型进行自我反思和设计批评的看法

Juan David Salazar Rodriguez, Sam Conrad Joyce, Nachamma Sockalingam, Khoo Eng Tat, Julfendi

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本研究探讨了大型语言模型在建筑工作室自我反思与设计批评中的应用,发现学生将其视为协作工具,帮助构建批判性思维并提升设计迭代效率。

Comments Keywords: Architectural Education, Design Studio Pedagogy, Large Lan-guage Models, Generative AI in Education, Design Critique

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17873 2026-02-04 cs.CV cs.AI 88%

SurgVidLM: Towards Multi-grained Surgical Video Understanding with Large Language Model

SurgVidLM:迈向多粒度外科视频理解的大型语言模型

Guankun Wang, Junyi Wang, Wenjin Mo, Long Bai, Kun Yuan, Ming Hu, Jinlin Wu, Junjun He, Yiming Huang, Nicolas Padoy, Zhen Lei, Hongbin Liu, Nassir Navab, Hongliang Ren

机构 * The Chinese University of Hong Kong(香港中文大学) Sun Yat-sen University(中山大学) University of Strasbourg(斯特拉斯堡大学) Technical University of Munich(慕尼黑技术大学) Monash University(墨尔本大学) Centre for Artificial Intelligence and Robotics, HKISI-CAS(人工智能与机器人中心,HKISI-CAS) Shanghai AI Laboratory(上海人工智能实验室)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 SurgVidLM通过多粒度分析提升外科视频理解能力,结合全局与局部机制实现更精确的手术流程解析。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02028 2026-01-30 cs.CL 88%

Are Large Language Models Good Classifiers? A Study on Edit Intent Classification in Scientific Document Revisions

大型语言模型是好的分类器吗?科学文档修订中的编辑意图分类研究

Qian Ruan, Ilia Kuznetsov, Iryna Gurevych

机构 * Ubiquitous Knowledge Processing Lab (UKP Lab)(通用知识处理实验室) Department of Computer Science(计算机科学系) Hessian Center for AI (hessian.AI)(黑森人工智能中心)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文研究了大型语言模型在科学文档修订中的编辑意图分类任务中的表现,通过实验和数据集构建探讨了其分类能力及应用价值。

Comments EMNLP2024 Main

Journal ref Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing EMNLP 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.19533 2026-01-28 cs.SD cs.AI 88%

SLM-SS: Speech Language Model for Generative Speech Separation

SLM-SS:用于生成式语音分离的语音语言模型

Tianhua Li, Chenda Li, Wei Wang, Xin Zhou, Xihui Chen, Jianqing Gao, Yanmin Qian

机构 * Auditory Cognition and Computational Acoustics Lab(听觉认知与计算声学实验室) MoE Key Lab of Artificial Intelligence, AI Institute(人工智能MoE重点实验室,AI研究院) School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院) VUI Labs(VUI实验室) AI research Institute, iFLYTEK Company Limited(AI研究院,iFLYTEK公司)

专题命中 领域大模型 :language model(title,abstract);SLM(title,abstract);分类 cs.AI

AI总结 SLM-SS通过应用语音语言模型提升生成式语音分离的语音可懂度和连贯性,实验显示其在语音可懂度和下游任务表现上优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10246 2026-01-28 cs.CL 88%

coTherapist: A Behavior-Aligned Small Language Model to Support Mental Healthcare Experts

coTherapist:一种行为对齐的小语言模型,用于支持心理健康专家

Prottay Kumar Adhikary, Reena Rawat, Tanmoy Chakraborty

机构 * IIT Delhi(德里印度理工学院)

专题命中 领域大模型 :language model(title,abstract);small language model(title,abstract);分类 cs.CL

AI总结 coTherapist通过小语言模型和领域微调等技术,为心理健康专家提供支持,展现出高同理心和临床可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15319 2026-01-23 q-bio.NC cs.AI 88%

Large Language Models as Simulative Agents for Neurodivergent Adult Psychometric Profiles

大语言模型作为神经多样性成人心理测量剖面的模拟代理

Francesco Chiappone, Davide Marocco, Nicola Milano

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本研究探讨大语言模型能否基于结构化访谈生成接近真实个体的心理测量反应,发现其在神经发育特征模拟中表现优异,但存在特定局限性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13388 2026-01-21 cs.CL 88%

Structured Insight from Unstructured Data: Large Language Models for SDOH-Driven Diabetes Risk Prediction

从无结构数据中提取结构化洞察:大型语言模型用于SDOH驱动的糖尿病风险预测

Sasha Ronaghi, Prerit Choudhary, David H Rehkopf, Bryant Lin

机构 * Stanford University(斯坦福大学) Stanford University School of Medicine(斯坦福大学医学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本研究利用大型语言模型从患者生活故事中提取结构化SDOH信息,用于糖尿病风险预测,LLMs在预测糖尿病控制水平方面达到60%的准确率。

Comments 7 pages, 5 figures

Journal ref Annu Int Conf IEEE Eng Med Biol Soc. 2025 Jul;2025:1-7

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17651 2026-01-21 cs.SE cs.AI 88%

Software Model Evolution with Large Language Models: Experiments on Simulated, Public, and Industrial Datasets

基于大语言模型的软件模型演化:在模拟、公开和工业数据集上的实验

Christof Tinnes, Alisa Welter, Sven Apel

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文提出RAMC方法,利用大语言模型实现软件模型补全,通过实验验证其在工业和公开数据集上的有效性。

Journal ref Proceedings of the 47th International Conference on Software Engineering (ICSE 2025), IEEE/ACM, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10167 2026-01-16 cs.CL 88%

Credit C-GPT: A Domain-Specialized Large Language Model for Conversational Understanding in Vietnamese Debt Collection

信用C-GPT:一种面向越南债务催收的领域专用大语言模型

Nhung Nguyen Thi Hong, Cuong Nguyen Dang, Tri Le Ngoc

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 Credit C-GPT是一种针对越南债务催收场景优化的领域专用大语言模型,通过整合多任务对话智能,提升对话理解、情感识别和意图检测等能力,为企业呼叫中心提供实时帮助和分析解决方案。

Comments 8 pages, 0 figures, 3 tables. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19562 2026-01-13 cs.AI 88%

FairMedQA: Benchmarking Bias in Large Language Models for Medical Question Answering

FairMedQA: 对大型语言模型在医学问答中的偏见进行基准测试

Ying Xiao, Jie Huang, Ruijuan He, Jing Xiao, Mohammad Reza Mousavi, Yepang Liu, Kezhi Li, Zhenpeng Chen, Jie M. Zhang

机构 * King’s College London(伦敦国王学院) University of Electronic Science and Technology of China(电子科技大学) Sun Yat-sen University(中山大学) York University(约克大学) Southern University of Science and Technology(南方科技大学) University College London(伦敦大学学院) Tsinghua University(清华大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 FairMedQA基准测试揭示了大型语言模型在医学问答中存在显著偏见,需进一步开发去偏技术和严格验证方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03988 2026-01-08 cs.SE cs.LG 88%

Using Small Language Models to Reverse-Engineer Machine Learning Pipelines Structures

利用小型语言模型反向工程机器学习流水线结构

Nicolas Lacroix, Mireille Blay-Fornarino, Sébastien Mosser, Frederic Precioso

机构 * Université Côte d’Azur, Inria, CNRS, I3S(法国蔚蓝海岸大学、Inria、CNRS、I3S)

专题命中 领域大模型 :language model(title,abstract);small language model(title,abstract);分类 cs.LG

AI总结 本文研究小型语言模型在反向工程机器学习流水线结构中的应用,评估其在提升数据科学实践理解方面的潜力。

Comments SANER 2026 Registered Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15834 2026-01-07 cs.CL cs.DL cs.IR q-bio.OT 88%

Scalable Scientific Interest Profiling Using Large Language Models

利用大语言模型实现可扩展的科学兴趣画像

Yilun Liang, Gongbo Zhang, Edward Sun, Betina Idnay, Yilu Fang, Fangyi Chen, Casey Ta, Yifan Peng, Chunhua Weng

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文利用大语言模型生成科学兴趣画像,通过MeSH术语和摘要两种方法对比,发现MeSH基于的画像在可读性和准确性上表现更优,但与人类写作存在概念选择差异。

Journal ref Journal of Biomedical Informatics 172, 104949 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06372 2026-01-06 cs.SD cs.AI 88%

SpeakerLM: End-to-End Versatile Speaker Diarization and Recognition with Multimodal Large Language Models

SpeakerLM:基于多模态大语言模型的端到端多功能说话人辨识与识别

Han Yin, Yafeng Chen, Chong Deng, Luyao Cheng, Hui Wang, Chao-Hong Tan, Qian Chen, Wen Wang, Xiangang Li

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 SpeakerLM通过多模态大语言模型实现端到端的说话人辨识与识别,结合灵活的说话人注册机制,提升多说话人场景下的识别性能。

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19228 2025-12-23 cs.AI 88%

Generation of Programmatic Rules for Document Forgery Detection Using Large Language Models

利用大语言模型生成文档伪造检测的程序规则

Valentin Schmidberger, Manuel Eberhardinger, Setareh Maghsudi, Johannes Maucher

机构 * Stuttgart Media University(斯图加特媒体大学) Ruhr-University Bochum(鲁尔大学波恩)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文利用大语言模型生成文档伪造检测的规则,通过微调策略在受限硬件上实现高效验证程序。

Comments Accepted at ICMLA 2025, the first two authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.16883 2025-12-19 cs.CL 88%

AdaSearch: Balancing Parametric Knowledge and Search in Large Language Models via Reinforcement Learning

AdaSearch:通过强化学习平衡参数知识与搜索

Tzu-Han Lin, Wei-Lin Chen, Chen-An Li, Hung-yi Lee, Yun-Nung Chen, Yu Meng

机构 * National Taiwan University(台湾大学) Department of Computer Science, University of Virginia(弗吉尼亚大学计算机科学系)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 AdaSearch通过强化学习框架平衡参数知识与搜索,提高知识边界意识并减少不必要的搜索调用。

Comments Preprint. Code and artifacts will be uploaded to https://github.com/hank0316/AdaSearch

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11544 2025-12-15 cs.AI 88%

AI-MASLD Metabolic Dysfunction and Information Steatosis of Large Language Models in Unstructured Clinical Narratives

AI-MASLD 大语言模型在无结构临床叙述中的代谢功能障碍与信息脂肪变

Yuan Shen, Xiaojun Wu, Linghua Yu

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本研究首次实证表明大语言模型在处理临床信息时表现出类似代谢功能障碍的特征,提出AI-MASLD概念,强调需在人类监督下使用LLMs作为辅助工具。

Comments 47 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10206 2025-12-15 cs.AI 88%

CP-Env: Evaluating Large Language Models on Clinical Pathways in a Controllable Hospital Environment

CP-Env:在可控医院环境中评估大型语言模型在临床路径中的表现

Yakun Zhu, Zhongzhen Huang, Qianhan Feng, Linjie Mu, Yannian Gu, Shaoting Zhang, Qi Dou, Xiaofan Zhang

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Innovation Institute(上海创新研究院) The Chinese University of Hong Kong(香港中文大学) SenseTime Research(商汤科技研究院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 CP-Env通过可控医院环境评估大型语言模型在复杂临床路径中的表现,揭示其在决策和伦理方面的挑战,并推动医疗AI的发展。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04228 2025-12-05 cs.AI 88%

Addressing Logical Fallacies In Scientific Reasoning From Large Language Models: Towards a Dual-Inference Training Framework

解决大型语言模型在科学推理中的逻辑谬误:一种双推理训练框架

Peter B. Walker, Hannah Davidson, Aiden Foster, Matthew Lienert, Thomas Pardue, Dale Russell

机构 * Intelligenesis LLC Uniformed Services University(美国武装部队服务大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文提出一种双推理训练框架,旨在解决大型语言模型在科学推理中的逻辑谬误问题,通过整合肯定生成与反事实否定,提升模型的鲁棒性和可解释性。

Comments 12 pages, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02024 2025-12-03 cs.CL cs.CY 88%

Human-Level and Beyond: Benchmarking Large Language Models Against Clinical Pharmacists in Prescription Review

人类水平与超越:在处方审查中对大型语言模型与临床药师进行基准测试

Yan Yang, Mouxiao Bian, Peiling Li, Bingjian Wen, Ruiyao Chen, Kangkun Mao, Xiaojun Ye, Tianbin Li, Pengcheng Chen, Bing Han, Jie Xu, Kaifeng Qiu, Junyan Wu

机构 * SUN YAT-SEN MEMORIAL HOSPITAL(中山纪念医院) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) University of Washington(华盛顿大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 RxBench通过系统评估大型语言模型在处方审查中的表现,揭示其能力与局限,并推动更可靠的临床工具开发。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01191 2025-12-02 cs.CL 88%

Generalist Large Language Models Outperform Clinical Tools on Medical Benchmarks

通用大语言模型在医疗基准测试中优于临床工具

Krithik Vishwanath, Mrigayu Ghosh, Anton Alyakin, Daniel Alexander Alber, Yindalon Aphinyanaphongs, Eric Karl Oermann

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 通用大语言模型在医疗基准测试中表现优于临床工具,揭示了临床AI系统在某些关键能力上的不足。

Comments 17 pages, 4 figures (2 regular, 2 supplemental)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23055 2025-12-01 cs.LG 88%

CDR-Agent: Intelligent Selection and Execution of Clinical Decision Rules Using Large Language Model Agents

CDR-Agent: 基于大语言模型代理的临床决策规则智能选择与执行

Zhen Xiang, Aliyah R. Hsu, Austin V. Zane, Aaron E. Kornblith, Margaret J. Lin-Martore, Jasmanpreet C. Kaur, Vasuda M. Dokiparthi, Bo Li, Bin Yu

机构 * University of Georgia(佐治亚大学) University of California, Berkeley(加州大学伯克利分校) University of California, San Francisco(加州大学旧金山分校) University of Chicago(芝加哥大学)

专题命中 领域大模型 :large language model(title);language model(title);LLM(abstract);prompting(abstract)

AI总结 CDR-Agent利用大语言模型代理,通过自主识别和应用最合适的临床决策规则,提升急诊部门的决策效率和准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏