arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12145 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12145 篇

2510.08284 2026-03-31 cs.CL 88%

Neuron-Level Analysis of Cultural Understanding in Large Language Models

大语言模型中文化理解的神经层面分析

Taisei Yamamoto, Ryoma Kumon, Danushka Bollegala, Hitomi Yanaka

机构 * The University of Tokyo(东京大学) Riken(理化学研究所) University of Liverpool(利物浦大学) Amazon(亚马逊) Tohoku University(东北大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 研究通过神经层面分析揭示大语言模型中文化理解的机制,发现文化通用与特定神经元对文化理解有重要影响,且训练数据影响其文化理解能力。

Comments Accepted to ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02572 2026-03-31 cs.CL 88%

Cultural Biases of Large Language Models and Humans in Historical Interpretation

大型语言模型和人类在历史解读中的文化偏见

Fabio Celli, Georgios Spathulas

机构 * Research & Development Department of Information Security and Communication Technology(信息安全和通信技术研发部) Maggioli SpA(Maggioli股份公司) Norwegian University of Science and Technology - NTNU(挪威科技大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文比较了人类和大型语言模型在历史注释中的表现,发现两者均存在文化偏见,但语言模型在短文本历史事实解读上达成更高共识,而人类因个人偏见常有分歧。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05181 2026-03-24 cs.CR cs.AI cs.CY 88%

Auditing Pay-Per-Token in Large Language Models

对大语言模型中按token计费的审计

Ander Artola Velasco, Stratis Tsirtsis, Manuel Gomez-Rodriguez

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文提出基于鞅理论的审计框架,用于检测大语言模型中的token误报问题,通过第三方审计者逐步查询服务提供商,确保能准确识别误报而避免误判。

Comments AISTATS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03637 2026-03-18 cs.AI 88%

Large Language Models for Combinatorial Optimization: A Systematic Review

大语言模型用于组合优化:系统综述

Francesca Da Ros, Michael Soprano, Luca Di Gaspero, Kevin Roitero

机构 * University of Udine(乌迪大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文系统回顾了大语言模型在组合优化中的应用,分析了103项研究,涵盖任务类型、模型架构、数据集及应用领域,并探讨了未来发展方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17618 2026-03-17 cs.CL cs.PF 88%

SimLens for Early Exit in Large Language Models: Eliciting Accurate Latent Predictions with One More Token

SimLens用于大型语言模型中的早期退出:通过一个额外的标记获取准确的潜在预测

Ming Ma, Bowen Zheng, Zhongqiao Lin, Tianming Yang

机构 * Institute of Neuroscience, State Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology, Center for Excellence in Brain Science and Intelligence Technology, Chinese Academy of Sciences(中国科学院脑科学与智能技术卓越创新中心神经科学研究所,脑认知与脑启发智能技术国家重点实验室,脑科学与智能技术卓越创新中心) School of Future Technology, University of Chinese Academy of Sciences(中国科学院大学未来技术学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文提出SimLens,一种无需训练的解码器,通过保留起始标记和候选答案标记进行轻量级延续,提升潜在预测准确性。结合线性SimLens和SimExit机制,在多个数据集上实现更高的准确性和更快的推理速度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04647 2026-03-06 cs.CL 88%

Coordinated Semantic Alignment and Evidence Constraints for Retrieval-Augmented Generation with Large Language Models

协同语义对齐与证据约束在大型语言模型检索增强生成中的应用

Xin Chen, Saili Uday Gadgil, Jiarong Qiu

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文提出一种通过协同语义对齐与证据约束提升检索增强生成效果的方法,增强事实可靠性和生成流畅性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04480 2026-03-06 q-bio.QM cs.LG 88%

AbAffinity: A Large Language Model for Predicting Antibody Binding Affinity against SARS-CoV-2

AbAffinity:一种用于预测抗SARS-CoV-2病毒抗体结合亲和力的大语言模型

Faisal Bin Ashraf, Animesh Ray, Stefano Lonardi

机构 * Department of Computer Science and Engineering,University of California, Riverside(加州大学河滨分校计算机科学与工程系) Riggs School of Applied Life Sciences, Keck Graduate Institute, Claremont(克劳斯研究生院应用生命科学学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 AbAffinity是一种通过大语言模型预测抗SARS-CoV-2病毒抗体结合亲和力的新型方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19297 2026-02-25 cs.AI 88%

Automated Generation of Microfluidic Netlists using Large Language Models

利用大语言模型自动化生成微流体网表

Jasper Davidson, Skylar Stockham, Allen Boston, Ashton Snelgrove, Valerio Tenace, Pierre-Emmanuel Gaillardon

机构 * Department of Electrical and Computer Engineering, University of Utah(电气与计算机工程系,犹他大学) Primis AI, Inc.(Primis AI 公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文利用大语言模型自动化生成微流体器件的结构网表,实现了从自然语言描述到Verilog代码的转换,并在典型微流体设计中验证了其有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13469 2026-02-20 cs.HC cs.AI 88%

How Multimodal Large Language Models Support Access to Visual Information: A Diary Study With Blind and Low Vision People

多模态大语言模型如何支持视障人士获取视觉信息:一项与盲人和低视力人士的日记研究

Ricardo E. Gonzalez Penuela, Crescentia Jung, Sharon Y Lin, Ruiying Hu, Shiri Azenkot

机构 * Cornell University(康奈尔大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本研究探讨了多模态大语言模型如何通过视觉助手技能支持视障人士获取视觉信息,并发现其在实际应用中的表现及改进方向。

Comments 24 pages, 17 figures, 7 tables, appendix section, to appear main track CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02178 2026-02-04 cs.CL 88%

AR-MAP: Are Autoregressive Large Language Models Implicit Teachers for Diffusion Large Language Models?

AR-MAP:自回归大语言模型是否是扩散大语言模型的隐式教师?

Liang Lin, Feng Xiong, Zengbin Wang, Kun Wang, Junhao Dong, Xuecai Hu, Yong Wang, Xiangxiang Chu

机构 * AMAP, Alibaba Group(AMAP,阿里巴巴集团) Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 AR-MAP通过利用自回归大语言模型作为隐式教师,有效提升扩散大语言模型的偏好对齐性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02132 2026-02-03 cs.CL 88%

There Is More to Refusal in Large Language Models than a Single Direction

大型语言模型中拒绝行为远不止单一方向

Faaiz Joad, Majd Hawasly, Sabri Boughorbel, Nadir Durrani, Husrev Taha Sencar

机构 * Qatar Computing Research Institute(卡塔尔计算研究所) HBKU(哈比卜大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 研究揭示大型语言模型中拒绝行为由多种几何方向控制,不同方向影响拒绝方式而非是否拒绝

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22020 2026-01-30 cs.LG cs.CV 88%

Visual-Guided Key-Token Regularization for Multimodal Large Language Model Unlearning

多模态大语言模型去敏中的视觉引导关键标记正则化

Chengyi Cai, Zesheng Ye, Peike Li, Bo Han, Jianzhong Qi, Feng Liu

机构 * The University of Melbourne(墨尔本大学) Google Research(谷歌研究) Hong Kong Baptist University(香港 Baptist 大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 本文提出视觉引导的关键标记正则化方法,用于多模态大语言模型的去敏,通过信息熵定义关键标记并利用梯度重新加权提升去敏效果,实验表明能有效减少遗忘并保持响应一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.16568 2026-01-28 cs.LG 88%

Predicting Startup Success Using Large Language Models: A Novel In-Context Learning Approach

利用大型语言模型预测初创企业成功:一种新颖的上下文学习方法

Abdurahman Maarouf, Alket Bakiaj, Stefan Feuerriegel

机构 * Munich Center for Machine Learning (MCML) & LMU Munich(慕尼黑机器学习中心(MCML)及慕尼黑大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 本文提出了一种基于k-最近邻的上下文学习方法,利用少量标注数据预测初创企业成功,证明其在数据稀缺环境下具有较高的预测准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15698 2026-01-23 cs.CV cs.AI 88%

Beyond Visual Safety: Jailbreaking Multimodal Large Language Models for Harmful Image Generation via Semantic-Agnostic Inputs

超越视觉安全:通过语义无关输入对多模态大语言模型进行有害图像生成的劫持

Mingyu Yu, Lana Liu, Zhehao Zhao, Wei Wang, Sujuan Qin

机构 * State Key Laboratory of Networking and Switching Technology, Beijing University of Posts and Telecommunications(网络与交换技术国家重点实验室,北京邮电大学) School of Cyberspace Security, Beijing University of Posts and Telecommunications(网络安全学院,北京邮电大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文提出BVS框架,通过语义无关输入对多模态大语言模型进行有害图像生成的劫持,揭示其视觉安全边界的脆弱性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12555 2026-01-21 cs.CL 88%

Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models

评估多语言大语言模型中的上下文中介事实回忆

Yihong Liu, Bingyu Xiong, Hinrich Schütze

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 研究多语言大语言模型在自然上下文中回忆事实的能力,发现上下文中介显著降低事实回忆准确性,大模型更鲁棒,真实名称影响不系统。

Comments preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08668 2026-01-14 cs.CL 88%

Analyzing Bias in False Refusal Behavior of Large Language Models for Hate Speech Detoxification

分析大型语言模型在仇恨言论净化中的虚假拒绝行为偏见

Kyuri Im, Shuzhou Yuan, Michael Färber

机构 * TU Dresden(德累斯顿理工大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 研究分析了大型语言模型在仇恨言论净化中的虚假拒绝偏见,并提出通过中英互译策略减少此类拒绝行为。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.13874 2026-01-14 cs.AI 88%

Geometry of Knowledge Allows Extending Diversity Boundaries of Large Language Models

知识的几何结构使大语言模型的多样性边界得以扩展

Mateusz Bystroński, Doheon Han, Nitesh V. Chawla, Tomasz Kajdanowicz

机构 * Wrocław University of Science and Technology(沃拉布大学科学与技术学院) University of Notre Dame(诺特丹大学)

专题命中 其他LLM :language model(title,abstract);large language model(title);LLM(abstract);分类 cs.AI

AI总结 基于知识的几何结构,通过流形条件调节扩展大语言模型的语义多样性边界,提升创造性发散思维。

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10808 2026-01-14 cs.CL 88%

ActiveLLM: Large Language Model-based Active Learning for Textual Few-Shot Scenarios

ActiveLLM: 基于大语言模型的文本少样本场景中的主动学习

Markus Bayer, Justin Lutz, Christian Reuter

机构 * PEASEC Technical University of Darmstadt(PEASEC技术大学达姆施塔特)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 ActiveLLM利用大语言模型提升少样本场景下的分类性能,优于传统方法及ADAPET、PERFECT和SetFit等少样本学习方法。

Comments 20 pages, 10 figures, 7 tables

Journal ref Transactions of the Association for Computational Linguistics 14 (2026) 1-22

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07347 2026-01-13 cs.CL 88%

DiffER: Diffusion Entity-Relation Modeling for Reversal Curse in Diffusion Large Language Models

DiffER: 用于扩散大语言模型中反转诅咒的扩散实体-关系建模

Shaokai He, Kaiwen Wei, Xinyi Zeng, Xiang Chen, Xue Yang, Zhenyang Li, Jiang Zhong, Yu Tian

机构 * Chongqing University(重庆大学) Tsinghua University(清华大学) Shanghai Jiao Tong University(上海交通大学) Nanjing University of Aeronautics and Astronautics(南京航空航天大学) Hong Kong University of Science and Technology(香港理工大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 DiffER通过实体感知训练和平衡数据构建,解决扩散大语言模型中的反转诅咒问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03831 2025-12-30 cs.LG physics.atm-clus physics.chem-ph physics.comp-ph 88%

A large language model-type architecture for high-dimensional molecular potential energy surfaces

用于高维分子势能面的大型语言模型型架构

Xiao Zhu, Srinivasan S. Iyengar

机构 * Department of Chemistry, Department of Physics(化学系、物理系) the Indiana University Quantum Science(印第安纳大学量子科学) Engineering Center (IU-QSEC), Indiana University, 800 E. Kirkwood Ave, Bloomington, IN-47405(工程中心(IU-QSEC),印第安纳大学,800 E. Kirkwood Ave,布卢明顿,IN-47405)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 本文提出了一种基于图神经网络的架构,用于高效计算高维分子势能面,实现了对186维势能面的高精度预测。

Comments 31 pages, 35 figures

Journal ref Phys. Rev. X, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.16063 2025-12-19 cs.HC cs.AI 88%

A Multi-Agent Large Language Model Framework for Automated Qualitative Analysis

一个多智能体大语言模型框架用于自动化定性分析

Qidi Xu, Nuzha Amjad, Grace Giles, Alexa Cumming, De'angelo Hermesky, Alexander Wen, Min Ji Kwak, Yejin Kim

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文提出CoTI框架,通过多智能体协作实现自动化定性分析,提升主题识别效率,但指出过度依赖AI可能影响研究者独立思考。

Comments 42 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13441 2025-12-17 cs.CL q-bio.NC 88%

Large language models are not about natural language

大语言模型并非关于自然语言

Johan J. Bolhuis, Andrea Moro, Stephen Crain, Sandiway Fong

机构 * University of Cambridge, Department of Psychology(剑桥大学心理学系) University School for Advanced Studies(高级研究大学) Scuola Normale Superiore(规范大学) Macquarie University, Department of Linguistics(麦考瑞大学语言学系) University of Arizona, Department of Linguistics(亚利桑那大学语言学系)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 大语言模型并非基于自然语言构建,而是基于概率模型,而人类语言由内在计算系统生成层次化思维结构。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24130 2025-12-16 cs.CL 88%

Beyond Magic Words: Sharpness-Aware Prompt Evolving for Robust Large Language Models with TARE

超越魔法词:为具有TARE的鲁棒大语言模型设计的敏锐性感知提示进化

Guancheng Wan, Lucheng Fu, Haoxin Liu, Yiqiao Jin, Hui Yi Leong, Eric Hanchen Jiang, Hejia Geng, Jinhe Bi, Yunpu Ma, Xiangru Tang, B. Aditya Prakash, Yizhou Sun, Wei Wang

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文提出TARE和ATARE方法,通过减少文本敏锐性差距,提升大语言模型在提示优化中的鲁棒性和准确性。

Comments We have identified a critical methodological error in Section 3 of the manuscript, which invalidates the main results; therefore, we request withdrawal for further revision

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.15781 2025-12-16 cs.CL 88%

DABL: Detecting Semantic Anomalies in Business Processes Using Large Language Models

基于大语言模型的业务流程语义异常检测:DABL

Wei Guan, Jian Cao, Jianqi Gao, Haiyan Zhao, Shiyou Qian

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 DABL利用大语言模型检测业务流程中的语义异常,通过生成正常轨迹和模拟异常,提升泛化能力和解释性。

Journal ref Proceedings of the AAAI Conference on Artificial Intelligence (AAAI), vol. 39, no. 11, pp. 11735-11744, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09117 2025-12-11 cs.AI 88%

A Categorical Analysis of Large Language Models and Why LLMs Circumvent the Symbol Grounding Problem

大型语言模型的范畴分析及为何LLMs绕过了符号 grounding 问题

Luciano Floridi, Yiyang Jia, Fernando Tohmé

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文通过范畴分析,探讨了LLMs如何绕过符号 grounding 问题,而非解决它。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07288 2025-12-09 cs.CL 88%

Investigating Training and Generalization in Faithful Self-Explanations of Large Language Models

探究大语言模型忠实自解释的训练与泛化

Tomoki Doi, Masaru Isonuma, Hitomi Yanaka

机构 * The University of Tokyo(东京大学) Riken(理化学研究所) Tohoku University(东北大学) NII LLMC(日本信息处理学会大语言模型委员会)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本研究通过训练提升大语言模型的自解释忠实性,并验证其在不同任务和风格中的泛化能力。

Comments To appear in the Proceedings of the Asia-Pacific Chapter of the Association for Computational Linguistics: Student Research Workshop (AACL-SRW 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00781 2025-12-09 cs.AI 88%

CoP: Agentic Red-teaming for Large Language Models using Composition of Principles

CoP: 用于大型语言模型的代理式红队测试:原理组合

Chen Xiong, Pin-Yu Chen, Tsung-Yi Ho

机构 * The Chinese University of Hong Kong Sha Tin, Hong Kong(香港中文大学(深圳)) IBM Research New York, USA(IBM纽约研究院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 CoP通过原理组合框架实现LLMs红队测试自动化,发现新型jailbreak提示并显著提升攻击成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22563 2025-12-03 cs.CL q-bio.NC 88%

Do Large Language Models Think Like the Brain? Sentence-Level Evidences from Layer-Wise Embeddings and fMRI

大语言模型是否像大脑思考?来自逐句嵌入和fMRI的层间证据

Yu Lei, Xingyang Ge, Yi Zhang, Yiming Yang, Bolei Ma

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本研究通过比较LLMs的层级嵌入与fMRI数据,揭示了大语言模型在句级层面与人类大脑的相似性,展示了LLMs在语言处理中的潜在应用价值。

Comments AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21097 2025-12-03 cs.SE cs.AI 88%

Model-Driven Quantum Code Generation Using Large Language Models and Retrieval-Augmented Generation

基于大语言模型和检索增强生成的模型驱动量子代码生成

Nazanin Siavash, Armin Moin

机构 * Department of Computer Science University of Colorado Colorado Springs (UCCS)(计算机科学系 佛罗里达大学科罗拉多州春分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文提出利用大语言模型和检索增强生成技术,通过UML模型生成量子代码,提升量子计算代码的准确性和一致性。

Comments This paper is accepted to the New Ideas and Emerging Results (NIER) track of the ACM/IEEE 28th International Conference on Model Driven Engineering Languages and Systems (MODELS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20494 2025-12-02 cs.CL 88%

Adversarial Confusion Attack: Disrupting Multimodal Large Language Models

对抗混淆攻击:破坏多模态大语言模型

Jakub Hoscilowicz, Artur Janicki

机构 * Warsaw University of Technology(华沙技术大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文提出对抗混淆攻击,通过生成扰动破坏多模态大语言模型的可靠性,展示其在不同模型上的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏