arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12581 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12581 篇

2508.10044 2025-12-02 cs.CR cs.AI 84%

Large Language Models for Power System Security: A Novel Multi-Modal Approach for Anomaly Detection in Energy Management Systems

用于电力系统安全的大型语言模型:一种用于能源管理系统异常检测的新型多模态方法

Aydin Zaboli, Junho Hong, Alexandru Stefanov, Chen-Ching Liu, Chul-Sang Hwang

机构 * Department of Electrical and Computer Engineering, University of Michigan -- Dearborn, MI, 48128 USA.(电气与计算机工程系,密歇根大学迪尔伯恩分校) Department of Electrical Sustainable Energy, Technische Universiteit Delft, 2628 CD Delft, Netherlands.(可持续能源系,代尔夫特理工大学) Bradley Department of Electrical and Computer Engineering, Virginia Polytechnic Institute and State University, Blacksburg, VA 24061, USA.(布雷德利电气与计算机工程系,弗吉尼亚理工学院和州立大学) Smart Grid Research Division System Reliability Research Team, Korea Electrotechnology Research Institute (KERI), Gwangju-si, 61751, South Korea.(智能电网研究分会系统可靠性研究团队,韩国电力技术研究所(KERI))

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.AI

AI总结 本文提出了一种基于大型语言模型的多模态方法,用于电力系统中能源管理系统的安全防护和异常检测。

Comments 10 Figures; 6 Tables; Accepted, IEEE ACCESS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.24119 2025-11-10 cs.SE cs.AI 84%

Leveraging Large Language Models for Code Translation and Software Development in Scientific Computing

Akash Dhruv, Anshu Dubey

机构 * Institute for Clarity in Documentation(清晰文档研究所) Inria Paris-Rocquencourt(巴黎-罗克琴特研究所) Rajiv Gandhi University(拉贾·甘地大学) Tsinghua University(清华大学) Palmer Research Laboratories(帕勒尔研究实验室) Argonne National Laboratory(阿贡国家实验室)

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04278 2025-08-07 cs.AI 84%

Large Language Model's Multi-Capability Alignment in Biomedical Domain

Wentao Wu, Linqing Chen, Hanmeng Zhong, Weilei Wang

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14162 2025-05-19 cs.AI 84%

EIAD: Explainable Industrial Anomaly Detection Via Multi-Modal Large Language Models

Zongyun Zhang, Jiacheng Ruan, Xian Gao, Ting Liu, Yuzhuo Fu

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.AI

Comments Accepted by ICME2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.16406 2025-03-04 cs.CL 84%

From RAGs to riches: Utilizing large language models to write documents for clinical trials

Nigel Markey, Ilyass El-Mansouri, Gaetan Rensonnet, Casper van Langen, Christoph Meier

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.CL

Comments 6 pages, 2 figures

Journal ref Clinical Trials: Journal of the Society for Clinical Trials (2025) -- online ahead of print

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05615 2025-02-11 cs.CV cs.AI 84%

XiHeFusion: Harnessing Large Language Models for Science Communication in Nuclear Fusion

Xiao Wang, Qingquan Yang, Fuling Wang, Qiang Chen, Wentao Wu, Yu Jin, Jingtao Jiang, Liye Jin, Bo Jiang, Dengdi Sun, Wanli Lv, Meiwen Chen, Zehua Chen, Guosheng Xu, Jin Tang

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08974 2025-01-16 cs.CL 84%

Learning to Extract Cross-Domain Aspects and Understanding Sentiments Using Large Language Models

Karukriti Kaushik Ghosh, Chiranjib Sur

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.07191 2025-01-14 eess.SY cs.LG cs.SY 84%

Pre-Trained Large Language Model Based Remaining Useful Life Transfer Prediction of Bearing

Laifa Tao, Zhengduo Zhao, Xuesong Wang, Bin Li, Wenchao Zhan, Xuanyuan Su, Shangyu Li, Qixuan Huang, Haifei Liu, Chen Lu, Zhixuan Lian

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.02346 2025-01-07 physics.med-ph cs.AI 84%

Exploring the Capabilities and Limitations of Large Language Models for Radiation Oncology Decision Support

Florian Putz, Marlen Haderleina, Sebastian Lettmaier, Sabine Semrau, Rainer Fietkau, Yixing Huang

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.AI

Comments Officially published in the Red Journal

Journal ref International Journal of Radiation Oncology, Biology, Physics. 2024 Mar 15;118(4):900-4

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05533 2024-12-10 cs.LG cs.CR 84%

Can large language models be privacy preserving and fair medical coders?

Ali Dadsetan, Dorsa Soleymani, Xijie Zeng, Frank Rudzicz

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.00175 2024-05-02 cs.CL cs.IR 84%

Towards a Search Engine for Machines: Unified Ranking for Multiple Retrieval-Augmented Large Language Models

Alireza Salemi, Hamed Zamani

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.16416 2024-01-26 cs.CL 84%

Improving Large Language Models for Clinical Named Entity Recognition via Prompt Engineering

Yan Hu, Qingyu Chen, Jingcheng Du, Xueqing Peng, Vipina Kuttichi Keloth, Xu Zuo, Yujia Zhou, Zehan Li, Xiaoqian Jiang, Zhiyong Lu, Kirk Roberts, Hua Xu

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.CL

Comments 17 pages, 5 tables, 6 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01040 2024-01-09 cs.CL 84%

From Beginner to Expert: Modeling Medical Knowledge into General LLMs

Qiang Li, Xiaoyan Yang, Haowen Wang, Qin Wang, Lei Liu, Junjie Wang, Yang Zhang, Mingyuan Chu, Sen Hu, Yicheng Chen, Yue Shen, Cong Fan, Wangshu Zhang, Teng Xu, Jinjie Gu, Jing Zheng, Guannan Zhang Ant Group

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);instruction tuning(abstract)

Comments Developed by Ant Group for PubMedQA leaderboard

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.02034 2024-01-05 cs.CL 84%

Text2MDT: Extracting Medical Decision Trees from Medical Texts

Wei Zhu, Wenfeng Li, Xing Tian, Pengfei Wang, Xiaoling Wang, Jin Chen, Yuanbin Wu, Yuan Ni, Guotong Xie

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);instruction tuning(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.11811 2023-11-21 cs.AI 84%

Large Language Models and Explainable Law: a Hybrid Methodology

Marco Billi, Alessandro Parenti, Giuseppe Pisano, Marco Sanchi

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.07197 2023-11-03 cond-mat.mtrl-sci cs.AI 84%

MatChat: A Large Language Model and Application Service Platform for Materials Science

Ziyi Chen, Fankai Xie, Meng Wan, Yang Yuan, Miao Liu, Zongguo Wang, Sheng Meng, Yangang Wang

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.AI

Journal ref Chinese Physics B 32, 118104 (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.17050 2023-10-02 cs.CL 84%

Interpretable Long-Form Legal Question Answering with Retrieval-Augmented Large Language Models

Antoine Louis, Gijs van Dijck, Gerasimos Spanakis

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.CL

Comments Under review. Code is available at https://github.com/maastrichtlawtech/lleqa

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.14522 2023-08-02 cs.CL 84%

CliniDigest: A Case Study in Large Language Model Based Large-Scale Summarization of Clinical Trial Descriptions

Renee D. White, Tristan Peng, Pann Sripitak, Alexander Rosenberg Johansen, Michael Snyder

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.CL

Comments 7 pages, 3 figures, 3 tables, conference: ACM GoodIt 23'; Second co-author: Tristan Peng; Citation: White, Peng, et al

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10250 2023-07-21 cs.AI 84%

Abductive Reasoning with the GPT-4 Language Model: Case studies from criminal investigation, medical practice, scientific research

Remo Pareschi

专题命中 领域大模型 :language model(title,abstract);large language model(abstract,comments);分类 cs.AI

Comments The article is 12 pages long and has one figure. It also includes a link to some ChatGPT dialogues that show the experiments that support the article's findings. The article will be published in V. Bambini and C. Barattieri di San Pietro (eds.), Sistemi Intelligenti, Special Section "Multidisciplinary perspectives on ChatGPT and the family of Large Language Models"

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.12135 2023-07-20 cs.CL 84%

Understand Legal Documents with Contextualized Large Language Models

Xin Jin, Yuchen Wang

专题命中 领域大模型 :large language model(title);language model(title);分类 cs.CL

Comments SemEval 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.16386 2026-08-18 cs.CL cs.LG 新提交 84%

Mint-Agent: Introducing Finance-Native Agentic Foundation Models

Mint-Agent:引入金融原生智能体基础模型

Agent Team, B. Zhang, Yaze Geng, Lei Tang, Yaoyang Yi, Zonghan Wu, Yifan Hu, Kun Wang, Qingsong Wen, Yilei Shao

专题命中 领域大模型 :foundation model(title);SFT(abstract,abstract_cn);分类 cs.CL、cs.LG

AI总结 本研究提出金融原生智能体模型Mint-Agent,通过三大支柱开发出Mint-Cu(9B)和Mint-Ag(27B),在多个金融基准测试中展现出优异的可靠性与可执行性,为可信金融智能提供了新路径。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.15254 2026-08-18 cs.AI cs.CL 新提交 84%

Demographic Injection in Medical Language Models under Diversity, Equity, and Inclusion Prompts

在多样性、公平性与包容性提示下医学语言模型的人口统计信息注入问题

Diego Mardian, Frank Liu

机构 * Arizona State University(亚利桑那州立大学)

专题命中 领域大模型 :language model(title,abstract);prompting(abstract);分类 cs.CL、cs.AI

AI总结 研究发现,在医学问题后附加DEI提示会使47个医学语言模型的人口统计信息注入率升至33.1%,部分虚构人口统计信息会误导模型输出错误答案,该效应源于公平性内容,是通用机制的体现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.14286 2026-08-17 cs.CV cs.AI cs.CL 新提交 84%

Seeing Red, Thinking Bad: Color Bias in Vision Language Models

看到红色,想到负面:视觉语言模型中的颜色偏差

Kohsuke Ide, Ryousuke Yamada, Yoshihiro Fukuhara, Hirokatsu Kataoka, Yutaka Satoh

机构 * National Institute of Advanced Industrial Science and Technology (AIST)(日本产业技术综合研究所) University of Tsukuba(筑波大学) University of Technology Nuremberg(纽伦堡工业大学) University of Oxford(牛津大学)

专题命中 领域大模型 :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本研究发现VLMs存在颜色偏差,通过引入Stealth Visual Prompts控制文本视觉风格,发现颜色、对比度会影响VLMs的情感预测和VQA输出,其行为与视觉编码器潜在表示变化相关。

Comments 15 pages. Accepted to ICPR 2026

Journal ref In: Pattern Recognition. ICPR 2026. Lecture Notes in Computer Science. Springer, Cham, pp. 261-275

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.10218 2026-08-12 cs.AI cs.CL 新提交 84%

Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems

思维病毒:多智能体大语言模型系统中的自传播思想

Vassilis Papadopoulos, McNair Shah, Sam Zimmerman, Jack Lindsey

专题命中 领域大模型 :LLM(title,summary_cn);分类 cs.CL、cs.AI

AI总结 本研究提出多智能体LLM系统中的思维病毒概念,构建其进化算法模型,发现其传播受宿主模型等因素影响,添加系统警告可免疫,为设计更稳健的多智能体系统提供依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09986 2026-08-05 cs.CL cs.AI 版本更新 84%

Quantifying Hallucinations in Language Language Models on Medical Textbooks

对语言模型在医学教科书中 hallucinations 的量化

Brandon C. Colelough, Davis Bartels, Dina Demner-Fushman

机构 * National Institutes of Health, National Library of Medicine(国家卫生研究院,国家医学图书馆) Department of Computer Science, University of Maryland(大学计算机科学系)

专题命中 领域大模型 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 本文研究了大型语言模型在医学教科书基础上的 hallucinations 发生频率及不同模型响应的差异,发现即使在高可信度响应下,LLaMA-70B-Instruct仍存在19.7%的hallucinations,且临床专家对模型响应的评估显示高一致性。

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.24743 2026-07-29 cs.CV cs.AI cs.CL 版本更新 84%

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding

ClinFusion:用于整体医学理解的以视觉为中心的多模态大语言模型系统

Hangjie Yuan, Yichen Qian, Zhiwei Tang, Xianzhe Xu, Lirong Wu, Sicheng Yang, Jinwang Wang, Pengju Wang, Zhitao Zeng, Yizeng Han, Yan Xing, Shengxuan Luo, Tao Feng, Qing Xie, Weigen Yao, Yi Yang, Zuozhu Liu, Jiasheng Tang, Shaocheng Wang, Jitao Wang, Jiahong Dong, Weihua Chen, Feng Xu, Fan Wang

机构 * DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团) Hupan Laboratory(湖畔实验室) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Department of Radiology, The Affiliated Yangming Hospital of Ningbo University(宁波大学附属阳明医院放射科) Zhejiang University-University of Illinois Urbana-Champaign Institute, Zhejiang University(浙江大学伊利诺伊大学厄巴纳香槟校区联合学院,浙江大学) Hepato-Pancreato-Biliary Center, Beijing Tsinghua Changgung Hospital, School of Clinical Medicine, Tsinghua Medicine, Tsinghua University(清华长庚医院肝胆胰中心,清华大学临床医学院,清华医学,清华大学) School of Software, Tsinghua University(清华大学软件学院) Beijing National Research Center for Information Science and Technology, Tsinghua University(清华大学北京信息科学与技术国家研究中心)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 研究针对多模态大语言模型在医学领域部署的挑战,提出ClinFusion,采用组合级联视觉编码器架构和视觉基础评估框架,在多模态医学基准测试中表现优异,超越开源和专有模型,经专家盲评验证效果良好。

Comments Code: https://github.com/alibaba-damo-academy/ClinFusion Models: https://huggingface.co/collections/Alibaba-DAMO-Academy/clinfusion

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.22552 2026-07-28 cs.CL cs.LG cs.SE 新提交 84%

MioFFAn: an Annotation Software for Formula Formalization with LLM Automation Capabilities

MioFFAn:一种具有大语言模型自动化能力的公式形式化注释软件

Nicolas Sibuet, Horacio Saggion, Riccardo Rossi

机构 * Universitat Politècnica de Catalunya (UPC)(加泰罗尼亚理工大学) Universitat Pompeu Fabra (UPF)(庞培法布拉大学) International Center for Numerical Methods in Engineering (CIMNE)(国际工程数值方法中心)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 针对科学文献数学表达式自动翻译受高质量数据集稀缺阻碍的问题,提出开源可定制框架MioFFAn,基于MioGatto架构扩展功能,通过大语言模型实现部分自动化,用标准NLP指标评估,初步证明人机协作方法有效。

Comments Presented in the 3rd International Workshop on Natural Scientific Language Processing (NSLP 2026), co-located at LREC2026

Journal ref Proceedings of the 3rd Int. Workshop on Natural Scientific Language Processing (NSLP 2026) at LREC 2026, pages 206-217

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.22100 2026-07-27 cs.CL cs.AI eess.AS 新提交 84%

MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond

MEUSLI:一种用于基于大语言模型的自动语音识别及其他应用的多语言投影器

Lorenzo Concina, Seraphina Fong, Marco Matassoni, Alessio Brutti

机构 * Center for Augmented Intelligence, Fondazione Bruno Kessler(增强智能中心,布鲁诺·凯斯勒基金会) Department of Information Engineering and Computer Science, University of Trento(特伦托大学信息工程与计算机科学系)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 MEUSLI是首个多语言投影器家族,连接Whisper编码器与开源多语言大语言模型,实现28种欧洲语言的开源端到端自动语音识别。它扩展了单语言管道,能借助持续学习技术扩展到其他语言,还可用于语音翻译和主题识别等任务。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20058 2026-07-23 cs.AI cond-mat.mes-hall cond-mat.mtrl-sci cs.CL 新提交 84%

Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model

在开放权重语言模型中读取和引导材料科学机制的表示

Markus J. Buehler

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 领域大模型 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 研究大语言模型中材料科学机制信息的表示形式,结合多种方法,包括匹配读数、状态几何等,通过实验验证其三种形式,如概念可读、取向由状态变换承载等,还通过比较提示等发现物理关系在受控状态变化中更易显现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.14385 2026-07-21 cs.CL cs.LG 版本更新 84%

MamaBench: Benchmarking LLM Robustness in Maternal and Child Health Diagnosis through Counterfactual Clinical Perturbation

MamaBench:通过反事实临床扰动对母婴健康诊断中的大语言模型稳健性进行基准测试

Thanni Adewuyi, Anuoluwa Sotome, Samuel Okoko, Angel Ezendu, Oluwafunke Akinbuwa, Oluwaseun Odunsi, Oluwasegun Oguntuase, Ifeoma Nwabueze, Abiodun Adereni

机构 * Helpmum Africa(非洲帮助妈妈组织) University of Ibadan(伊巴丹大学)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 研究针对大语言模型在母婴健康诊断中缺乏对临床相似表现区分能力的问题,提出MamaBench基准及EA - RAG方法,通过实验揭示基础准确率高估稳健准确率现象,EA - RAG有效降低BTR,提升稳健准确率,为临床人工智能反事实稳健性研究提供参考。

详情

展开后加载摘要…

URL PDF HTML 收藏