arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12496 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12496 篇

2512.14732 2026-08-17 cs.LG cs.AI cs.CV eess.IV 版本更新 90%

INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT

INFORM-CT:整合LLM和VLM用于腹部CT的偶发发现管理

Idan Tankel, Nir Mazor, Rafi Brada, Christina LeBedis, Guy ben-Yosef

机构 * GE Healthcare Technology and Innovation Center(GE医疗技术与创新中心) Boston Medical Center(波士顿医疗中心)

专题命中 领域大模型 :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于LLM和VLM的计划-执行框架,用于提高腹部CT偶发发现的检测、分类和报告效率与精度,通过自动化流程提升临床应用效果。

Comments Spotlight presentation at the 9th International Conference on Medical Imaging with Deep Learning (MIDL) 2026 Additional code and implementation details available at this https URL (https://idan-tankel.github.io/InformCT_ProjectPage/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01826 2026-08-17 cs.CL cs.AI 版本更新 90%

Leveraging Few-Shot Learning and Large Language Models for Analyzing Blood Pressure Variations Across Biological Sex from Scientific Literature

利用小样本学习与大语言模型从科学文献中分析不同生物性别间的血压差异

Yuting Guo, Seyedeh Somayyeh Mousavi, Reza Sameni, Abeed Sarker

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本研究利用小样本学习、LLaMA3及GPT-3.5等大语言模型,从PubMed文献中提取血压相关信息,分析不同生物性别间的血压差异,生成可视化图表开展研究。

Comments Accepted by the journal of Computers in Biology and Medicine

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.09548 2026-08-12 cs.CL cs.AI cs.CY 版本更新 90%

ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language Models

ELBench:面向教育场景的大语言模型多维基准

Yilin Jiang, Xiaorong Zhu, Fei Tan, Zicheng Zhang, Kaiyi Huang, Yang Yu, Zexuan Fei, Yiming Luo, Keqian Li, Hao Hao, Guangtao Zhai, Aimin Zhou

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);post-training(abstract);分类 cs.CL、cs.AI

AI总结 ELBench是首个评估教育大模型四项核心要求的综合基准,评估发现通用模型综合表现相近但模块优势不同,中国模型在安全性模块领先,教育专用模型未在教育模块占优,高阶培养存在系统性盲点。

Comments 13 pages, 6 figures, 8 tables. Benchmark data: https://huggingface.co/datasets/ZeroLoss-Lab/ELBench

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.28098 2026-07-31 cs.AI cs.CL 新提交 90%

SciDataSailor: Deep Scientific Data Exploring

SciDataSailor:深度科学数据探索

Jiyong Rao, Yicheng Qiu, Chi Zhang, Chunfeng Song, Runkai Zhao

专题命中 领域大模型 :LLM(summary_cn,abstract);SFT(abstract,abstract_cn);large language model(abstract);language model(abstract)

AI总结 该研究针对科学数据仓库交互难题,提出SciDataSailor框架,以带特定机制的MCTS实现轨迹合成,构建了微调模型与含千余任务的评估基准,推动LLM智能体的科学数据探索能力。

Comments 63 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26670 2026-07-30 cs.DL cs.AI cs.CL cs.IR 新提交 90%

Scientific Knowledge Discovery in the Age of Large Language Models

大语言模型时代的科学知识发现

Eleni Adamidi, Serafeim Chatzopoulos, Thanasis Vergoulis

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL、cs.AI

AI总结 本章综述34篇应用生成式大语言模型的同行评审论文,针对学术文献检索与合格研究筛选任务,基于OpenAIRE Graph布尔搜索筛选文献,为科学知识发现提供更灵活方案。

Comments 21 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12393 2026-07-27 cs.CL cs.AI 版本更新 90%

MedKGent: A Large Language Model Agent Framework for Constructing Temporally Evolving Medical Knowledge Graph

MedKGent:用于构建随时间演变的医学知识图谱的大语言模型智能体框架

Duzhen Zhang, Zixiao Wang, Zhong-Zhi Li, Yahan Yu, Shuncheng Jia, Jiahua Dong, Haotian Xu, Xing Wu, Yingying Zhang, Tielin Zhang, Jie Yang, Xiuying Chen, Le Song

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德大学人工智能学院) University of Chinese Academy of Sciences(中国科学院大学) Kyoto University(京都大学) Tsinghua University(清华大学) East China Normal University(华东师范大学) Center for Excellence in Brain Science and Intelligence Technology(脑科学与智能技术卓越中心) Brigham and Women’s Hospital, Harvard Medical School(哈佛医学院布里特妇女医院) GenBio AI

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 研究针对医学文献增长带来的知识结构化挑战,引入MedKGent框架,利用PubMed摘要通过两个智能体每日增量构建医学知识图谱,经评估其三元组有效性高,能显著改进大语言模型在医学问答基准上的检索增强生成。

Comments Accepted by Npj Digital Medicine

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.12336 2026-07-15 cs.CL cs.AI cs.CY cs.ET cs.HC 新提交 90%

Evaluating Health Misinformation in Low-Resource Languages: Integrating Small Language Models with a Culturally-Sensitive Responsible NLP Framework (Bangla as a Case Study)

评估低资源语言中的健康错误信息:将小语言模型与文化敏感的负责任自然语言处理框架相结合(以孟加拉语为例)

Farnaz Farid, Raihan Alam, Al Al-Areqi, Farhad Ahamed, Muhammad Hassan Khan, Sadia Hossain, Irena Veljanova, Anika Tabassum Binte Hossain

机构 * Western Sydney University(西悉尼大学) Microsoft(微软公司) Excelsia College(埃克塞尔西亚学院) Faulconbridge Health Centre(福尔康布里奇健康中心)

专题命中 领域大模型 :language model(title,abstract);small language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 研究针对低资源语言中健康错误信息难检测问题,提出结合小语言模型与文化敏感的负责任自然语言处理框架,以孟加拉语为例进行实验,证明Phi-4表现优,还设计新框架,为评估低资源语言错误信息提供整体视角。

Comments 39 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06078 2026-06-26 cs.CY cs.AI cs.CL 90%

Simulating Students with Large Language Models: A Review of Architecture, Mechanisms, and Role Modelling in Education with Generative AI

用大型语言模型模拟学生:生成式AI在教育中的架构、机制与角色建模综述

Luis Marquez-Carpintero, Alberto Lopez-Sellers, Miguel Cazorla

机构 * Institute for Computer Research University of Alicante(计算机研究所阿利坎特大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文综述了利用大型语言模型模拟学生行为在教育中的应用,探讨了其在学习者建模、教学评估和教师培训中的潜力与挑战。

Journal ref Computer Science Review 62 (2026) 101008

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20929 2026-06-23 cs.CL cs.AI 新提交 90%

Peeking Inside LLMs: Leveraging Internal Artifacts of LLMs for Enhancing Reliability in Legal Classification

窥视LLM内部:利用LLM的内部构件增强法律分类的可靠性

Sudipta Santra, Debtanu Datta, Saptarshi Ghosh

机构 * Indian Institute of Technology Kharagpur(印度理工学院卡哈拉格普尔分校)

专题命中 领域大模型 :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 针对LLM在法律领域易产生错误或幻觉的问题,提出利用模型内部构件特征构建下游分类器检测预测正确性,在保释决策和法规违规预测任务上验证了有效性。

Comments Accepted at the International Workshop on Automated Semantic Analysis of Information in Law (ASAIL) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15314 2026-06-16 cs.LG cs.AI stat.ML 新提交 90%

LLMs on Tabular Data with Limited Semantics: Evidence from Industrial Car Retrofit Prediction

有限语义表格数据上的LLM:来自工业汽车改造预测的证据

Aina Vila Pons, Ioannis Tzachristas, Constantinos Antoniou

机构 * Technical University of Munich(慕尼黑工业大学) BMW Group(宝马集团)

专题命中 领域大模型 :LLM(title_cn,summary_cn);foundation model(abstract);prompting(abstract);分类 cs.AI、cs.LG

AI总结 研究在工业表格数据中,LLM(嵌入、直接分类、混合堆叠)与经典树集成方法的对比,发现LLM在语义受限时效果有限,但嵌入和混合方法仍有价值。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14892 2026-06-15 cs.LG cs.AI 版本更新 90%

Can LLMs Accurately Score Medical Diagnoses and Clinical Reasoning?

LLM能否准确评分医学诊断和临床推理?

Amy Rouillard, Sitwala Mundia, Linda Camara, Ziyaad Dangor, Michael Cameron Gramanie, Ismail Kalla, Shabir A. Madhi, Kajal Morar, Marlvin T. Ncube, Haroon Saloojee, Bruce A. Bassett

机构 * Wits MIND Institute, University of the Witwatersrand, Johannesburg, South Africa(维特士心理研究所,沃斯兰德大学,约翰内斯堡,南非) Grai Labs, Cape Town, South Africa(格雷实验室,开普敦,南非) South African Medical Research Council Vaccines and Infectious Diseases Analytics Research Unit, Faculty of Health Sciences, University of the Witwatersrand, Johannesburg, South Africa(南非医学研究理事会疫苗和传染病分析研究组,健康科学学院,沃斯兰德大学,约翰内斯堡,南非) Department of Internal Medicine, Charlotte Maxeke Johannesburg Academic Hospital, and Faculty of Health Sciences, University of the Witwatersrand, Johannesburg, South Africa(内科学系,查理·马克斯凯约翰内斯堡学术医院,以及健康科学学院,沃斯兰德大学,约翰内斯堡,南非) Department of Paediatrics and Child Health, Faculty of Health Sciences, University of the Witwatersrand, Johannesburg, South Africa(儿科学与儿童健康系,健康科学学院,沃斯兰德大学,约翰内斯堡,南非) Wits MIND Institute, University of the Witwatersrand, Johannesbu(维特士心理研究所,沃斯兰德大学,约翰内斯堡)

专题命中 领域大模型 :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 研究使用LLM陪审团对300例低收入和中等收入国家医院病例的3334个诊断进行评分,发现校准后的LLM评分与专家评分高度一致,且严重错误风险更低,可作为可靠的评估代理。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.15467 2026-05-18 cs.CL cs.AI 90%

Retrieval-Augmented Large Language Models for Schema-Constrained Clinical Information Extraction

基于检索增强的大型语言模型用于受模式约束的临床信息提取

A H M Rezaul Karim, Ozlem Uzuner

机构 * George Mason University(乔治·马歇尔大学)

专题命中 领域大模型 :large language model(title);language model(title);LLM(abstract,abstract_cn);prompting(abstract)

AI总结 本文提出一种模块化检索增强生成框架,通过schema约束提示、确定性后处理和二次审核,提升护士-患者对话中观察提取的F1分数达80.36%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23430 2026-04-28 cs.IR cs.AI cs.CL cs.DL cs.SE 90%

Automating Categorization of Scientific Texts with In-Context Learning and Prompt-Chaining in Large Language Models

利用大语言模型中的上下文学习和提示链自动分类科学文本

Gautam Kishore Shahi, Oliver Hummel

机构 * Technische Hochschule Mannheim(曼海姆技术大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract_cn);分类 cs.CL、cs.AI

AI总结 本文评估了大语言模型在科学文本分类中的性能,发现提示链在处理ORKG层级分类时优于纯上下文学习,但在第三层级主题分类上仍存在50%的准确率限制。

Comments 25 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.28258 2026-04-27 cs.CL cs.AI 90%

Categorical Perception in Large Language Model Hidden States: Structural Warping at Digit-Count Boundaries

大语言模型隐藏状态中的分类感知:数字计数边界处的结构扭曲

Jon-Paul Cacioli

机构 * Independent Researcher(独立研究者) Classical Minds, Modern Machines(经典思维,现代机器)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract_cn);分类 cs.CL、cs.AI

AI总结 研究发现大语言模型在处理阿拉伯数字时,隐藏状态表现出类似人类分类感知的几何扭曲现象,通过代表相似性分析发现CP-加法模型在结构边界处更符合几何结构。

Comments 25 pages, 5 figures, 7 tables. Pre-registered on OSF (osf.io/qrxf3). Code at https://github.com/synthiumjp/weber

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.20791 2026-04-23 cs.CL cs.AI 90%

Can "AI" Be a Doctor? A Study of Empathy, Readability, and Alignment in Clinical LLMs

AI能成为医生吗?对临床LLM中共情、可读性和对齐性的研究

Mariano Barone, Francesco Di Serio, Roberto Moio, Marco Postiglione, Giuseppe Riccio, Antonio Romano, Vincenzo Moscato

机构 * Department of Electrical Engineering and Information Technology(电子工程与信息科技系) University of Naples Federico II(那不勒斯费德里科二世大学) Department of Translational Medical Sciences(转化医学科学系) University of Campania ”Luigi Vanvitelli”(坎帕尼亚“路易吉·范维蒂利”大学) Department of Computer Science, McCormick School of Engineering and Applied Science(计算机科学系,麦科马克工程与应用科学学院) Northwestern University(西北大学)

专题命中 领域大模型 :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文研究了临床LLM在共情、可读性和对齐性方面的表现,发现协作重写能显著提升对齐效果,但LLM仍无法超越医生的临床专业性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14179 2026-04-17 cs.CL cs.AI 90%

An Underexplored Frontier: Large Language Models for Rare Disease Patient Education and Communication -- A scoping review

未被充分探索的前沿:大型语言模型在罕见病患者教育与沟通中的应用——一项范围综述

Zaifu Zhan, Yu Hou, Kai Yu, Min Zeng, Anita Burgun, Xiaoyi Chen, Rui Zhang

机构 * Division of Computational Health Sciences, Department of Surgery, University of Minnesota(计算健康科学系,外科部,明尼苏达大学) Department of Electrical and Computer Engineering, University of Minnesota(电气与计算机工程系,明尼苏达大学) Clinical Bioinformatics Lab, Université Paris Cité, Institut Imagine, INSERM UMR 1163(临床生物信息学实验室,巴黎城市大学,Imagine机构,INSERM UMR 1163) Medical Informatics Department, Hôpital Necker-Enfants Malades, AP-HP(医学信息学部,Necker-Enfants Malades医院,AP-HP)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文通过范围综述探讨大型语言模型在罕见病患者教育与沟通中的应用,发现现有研究多集中于通用模型,缺乏真实场景和患者中心设计。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06737 2026-04-13 cs.CL cs.AI 90%

WisdomInterrogatory (LuWen): An Open-Source Legal Large Language Model Technical Report

WisdomInterrogatory (LuWen): 一种开源法律大语言模型技术报告

Yiquan Wu, Yuhang Liu, Yifei Liu, Ang Li, Siying Zhou, Kun Kuang, Fei Wu

机构 * Zhejiang University(浙江大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);foundation model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出基于Baichuan模型的开源中文法律语言模型LuWen,通过持续预训练、监督微调和检索增强生成技术,提升法律领域任务表现。

Comments 10 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08333 2026-04-10 cs.CV cs.AI cs.LG 90%

Lost in the Hype: Revealing and Dissecting the Performance Degradation of Medical Multimodal Large Language Models in Image Classification

迷失在喧嚣中:揭示并剖析医学多模态大语言模型在图像分类中的性能退化

Xun Zhu, Fanbin Mo, Xi Chen, Kaili Zheng, Shaoshuai Yang, Yiming Shi, Jian Gao, Miao Li, Ji Wu

机构 * College of AI, Tsinghua University(清华大学人工智能学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

AI总结 本文研究了医学多模态大语言模型在图像分类中的性能退化问题,通过实验和特征探查揭示了四个失败模式,并提出了量化评分用于评估特征演变的健康状况。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06028 2026-04-08 cs.CL cs.AI cs.IR 90%

A Multi-Stage Validation Framework for Trustworthy Large-scale Clinical Information Extraction using Large Language Models

一种用于可信大规模临床信息提取的多阶段验证框架

Maria Mahbub, Gregory M. Dams, Josh Arnold, Caitlin Rizy, Sudarshan Srinivasan, Elliot M. Fielstein, Minu A. Aghevli, Kamonica L. Craig, Elizabeth M. Oliva, Joseph Erdos, Jodie Trafton, Ioana Danciu

机构 * Oak Ridge National Laboratory(橡树岭国家实验室) Program Evaluation and Resource Center, Office of Mental Health and Office of Suicide Prevention, Department of Veterans Affairs(退伍军人事务部心理健康办公室和自杀预防办公室项目评估与资源中心) Vanderbilt University Medical Center(范德比尔特大学医学中心) VA Maryland Health Care System(退伍军人事务部马里兰医疗保健系统) VA Desert Pacific Healthcare Network(退伍军人事务部沙漠太平洋医疗网络) VA Connecticut Health Care System(退伍军人事务部康涅狄格医疗保健系统) Yale School of Medicine(耶鲁医学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文提出多阶段验证框架,通过弱监督评估提升LLM在临床信息提取中的可信度与准确性,验证了在大规模数据中应用的可行性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.01538 2026-04-03 cs.CL cs.AI 90%

Countering Catastrophic Forgetting of Large Language Models for Better Instruction Following via Weight-Space Model Merging

通过权重空间模型融合提升大语言模型的指令跟随能力以对抗灾难性遗忘

Mengxian Lyu, Cheng Peng, Ziyi Chen, Mengyuan Zhang, Jieting Li Lu, Yonghui Wu

机构 * University of Florida(佛罗里达大学) Department of Health Outcomes and Biomedical Informatics, College of Medicine, University of Florida(佛罗里达大学医学院健康结果与生物医学信息学系) Department of Engineering Education, Herbert Wertheim College of Engineering, University of Florida(佛罗里达大学赫伯特·韦特海姆工程学院工程教育系) Preston A. Wells, Jr. Center for Brain Tumor Therapy, Lillian S. Wells Department of Neurosurgery, University of Florida(佛罗里达大学莉莉安·S·威尔斯神经外科系普雷斯顿·A·威尔斯脑肿瘤治疗中心)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);foundation model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出通过权重空间模型融合方法,有效缓解大语言模型在医疗领域微调时的灾难性遗忘问题,保留指令跟随能力并提升医疗任务表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08592 2026-04-02 cs.CL cs.AI 90%

The Collective Turing Test: Large Language Models Can Generate Realistic Multi-User Discussions

集体图灵测试:大型语言模型可以生成逼真的多用户讨论

Azza Bouleimen, Giordano De Marzo, Taehee Kim, Nicol`o Pagan, Hannah Metzler, Silvia Giordano, Anikó Hannák, David Garcia

机构 * University of Zurich(苏黎世大学) University of Konstanz(康斯坦茨大学) Complexity Science Hub(复杂性科学中心) University of Applied Sciences and Arts of Southern Switzerland(瑞士南部应用科学与艺术大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文研究了大型语言模型在模拟社交媒体多用户讨论中的真实性,通过对比人类和AI生成的讨论内容,发现AI生成的讨论有39%被误认为真实,揭示了LLM在社交模拟中的潜力与潜在滥用风险。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.27164 2026-03-31 cs.AI cs.CL 90%

daVinci-LLM:Towards the Science of Pretraining

daVinci-LLM:迈向预训练的科学

Yiwei Qin, Yixiu Liu, Tiantian Mi, Muhang Xie, Zhen Huang, Weiye Si, Pengrui Lu, Siyuan Feng, Xia Wu, Liming Liu, Ye Luo, Jinlong Hou, Qipeng Guo, Yu Qiao, Pengfei Liu

机构 * SII SJTU(上海交通大学) GAIR

专题命中 领域大模型 :LLM(title,abstract);pretraining(title,abstract);post-training(abstract);分类 cs.CL、cs.AI

AI总结 本文提出daVinci-LLM,通过结合工业级资源与科研自由,探索预训练的科学方法。采用开放范式,通过数据达尔文主义框架和两阶段适应课程,训练3B参数模型,验证处理深度对能力的影响及不同领域饱和动态的差异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.23515 2026-03-26 cs.CL cs.AI 90%

Training a Large Language Model for Medical Coding Using Privacy-Preserving Synthetic Clinical Data

利用隐私保护的合成临床数据训练大型语言模型进行医学编码

John Cook, Michael Wyatt, Peng Wei, Iris Chin, Santosh Gupta, Van Zyl Van Vuuren, Richie Siburian, Amanda Spicer, Kristen Viviano, Alda Cami, Raunaq Malhotra, Zhewei Yao, Jeff Rasley, Gaurav Kaushik

机构 * Veradigm Snowflake

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);foundation model(abstract);分类 cs.CL、cs.AI

AI总结 本文研究了利用隐私保护的合成临床数据训练大型语言模型进行医学编码的可行性,通过微调Llama 3-70B模型,实现了ICD-10-CM和CPT代码的高精度预测,显著提升了编码准确性。

Comments 20 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.19271 2026-03-23 cs.CL cs.AI 90%

A Human-Centered Workflow for Using Large Language Models in Content Analysis

面向人类的大型语言模型在内容分析中的工作流程

Ivan Zupic

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种面向人类的大型语言模型在内容分析中的工作流程,涵盖注释、摘要和信息提取三个任务,强调通过API利用LLM的潜力,并提供验证方法和最佳实践以解决黑盒、提示敏感和幻觉等问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21599 2026-03-13 cs.IR cs.AI cs.LG 90%

Refine-POI: Reinforcement Fine-Tuned Large Language Models for Next Point-of-Interest Recommendation

Refine-POI: 通过强化微调优化大型语言模型用于下一步兴趣点推荐

Peibo Li, Shuang Ao, Hao Xue, Yang Song, Maarten de Rijke, Johan Barthélemy, Tomasz Bednarz, Flora D. Salim

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);SFT(abstract);分类 cs.AI、cs.LG

AI总结 Refine-POI通过拓扑感知ID生成和强化微调,解决LLMs在POI推荐中的语义连续性和top-k推荐问题,提升推荐准确性和可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00076 2026-03-03 cs.CY cs.AI cs.LG 90%

The Value Sensitivity Gap: How Clinical Large Language Models Respond to Patient Preference Statements in Shared Decision-Making

价值敏感度缺口:临床大语言模型在共享决策中的患者偏好响应

Sanjay Basu

机构 * Waymark University of California, San Francisco(加州大学旧金山分校)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

AI总结 本文研究了临床大语言模型在共享决策中对患者偏好声明的响应,发现不同模型在价值敏感度和一致性上存在差异,通过实验验证了模型对患者价值的识别能力。

Comments 38 pages, 4 figures, supplementary appendix included

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22215 2026-02-27 cs.AI cs.CL cs.IR 90%

Graph Your Way to Inspiration: Integrating Co-Author Graphs with Retrieval-Augmented Generation for Large Language Model Based Scientific Idea Generation

通过图谱获取灵感:将合著者图谱与检索增强生成结合用于基于大语言模型的科学想法生成

Pengzhen Xie, Huizhi Liang

机构 * School of Computing Newcastle University(计算学院新castle大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文提出GYWI系统,结合作者图谱与RAG技术,提升大语言模型生成科学想法的可控性和灵感追溯能力。

Comments 15 pages, 10 figures. Submitted to [RAAI]

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19139 2026-02-26 cs.AI cs.CL 90%

A Multi-faceted Analysis of Cognitive Abilities: Evaluating Prompt Methods with Large Language Models on the CONSORT Checklist

对认知能力的多维度分析:利用大型语言模型在CONSORT清单上评估提示方法

Sohyeon Jeon, Hyung-Chul Lee

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本研究通过比较通用和领域专用LLM在三种提示策略下的表现,揭示了其在评估临床试验报告时存在显著的校准错误和过度自信问题,强调了改进校准和提示工程的重要性。

Comments We have decided to withdraw this manuscript because we believe it requires further revision and substantial improvement before it is suitable for dissemination to the academic community

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15689 2026-02-19 cs.CL cs.AI cs.CR 90%

A Content-Based Framework for Cybersecurity Refusal Decisions in Large Language Models

基于内容的大型语言模型网络安全拒绝决策框架

Noa Linder, Meirav Segal, Omer Antverg, Gil Gekker, Tomer Fichman, Omri Bodenheimer, Edan Maor, Omer Nevo

机构 * University of Zurich(苏黎世大学) Irregular

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文提出基于内容的网络安全拒绝决策框架,通过显式建模进攻风险与防御效益的权衡,解决现有方法在一致性、过度限制和脆弱性方面的不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15791 2026-02-18 cs.AI cs.CL 90%

Enhancing Building Semantics Preservation in AI Model Training with Large Language Model Encodings

利用大语言模型编码增强AI模型训练中的建筑语义保持

Suhyung Jang, Ghang Lee, Jaekun Lee, Hyunjun Lee

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本研究利用大语言模型编码提升AI对建筑语义的理解,通过实验验证其在分类任务中的优越性。

Comments 42nd International Symposium on Automation and Robotics in Construction (ISARC 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏