arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7539 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7539 篇

2204.03489 2022-04-08 cs.CL cs.LG 84%

Position-based Prompting for Health Outcome Generation

M. Abaho, D. Bollegala, P. Williamson, S. Dodd

专题命中 知识编辑与模型理解 :prompting(title,abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.02647 2022-02-08 cs.CL cs.AI 84%

Ethics, Rules of Engagement, and AI: Neural Narrative Mapping Using Large Transformer Language Models

Philip Feldman, Aaron Dant, David Rosenbluth

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments 18 Pages, 13 figures

Journal ref Bulletin of the Technical Committee on Data Engineering, Vol. 44 No. 4 December 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.00781 2020-07-22 cs.CL cs.LG stat.ML 84%

On the comparability of Pre-trained Language Models

Matthias Aßenmacher, Christian Heumann

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.LG

Journal ref Proceedings of the 5th Swiss Text Analytics Conference (SwissText) & 16th Conference on Natural Language Processing (KONVENS), Zurich, Switzerland, June 23-25, 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08744 2024-09-16 cs.CV cs.LG 84%

Uncertainty and Generalizability in Foundation Models for Earth Observation

Raul Ramos-Pollan, Freddie Kalaitzis, Karthick Panner Selvam

专题命中 知识编辑与模型理解 :foundation model(title,abstract);pretraining(abstract);分类 cs.LG

Comments A large ablation study measuring uncertainty and spatial generalizability with 8 foundation models, 11 world regions and 7 downstream tasks

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.13900 2026-08-17 cs.DB cs.AI cs.CL cs.LG 新提交 83%

Agentic Transaction: Towards ACID-Compliant Agent Systems

智能体事务:面向ACID兼容的智能体系统

Zhaoyan Sun, Xiaoxiao Wang, Guoliang Li

专题命中 知识编辑与模型理解 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 该研究提出ACID兼容的智能体事务框架,开发对应数据智能体,在基准测试中较含Claude Code的现有智能体提升10.6%,为构建可信可扩展AI智能体开辟新方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16616 2026-08-13 cs.DB 版本更新 83%

LDI: Localized Data Imputation for Text-Rich Tables

LDI:面向文本丰富表的局部数据填补

Soroush Omidvartehrani, Davood Rafiei

专题命中 知识编辑与模型理解 :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract)

AI总结 本文提出LDI框架,通过局部推理利用LLM填补文本丰富表中的缺失值,提升准确性和可解释性,实验证明其优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04765 2026-05-29 cs.CL cs.AI cs.LG physics.comp-ph 83%

Differential syntactic and semantic encoding in LLMs

大型语言模型中句法与语义的差异编码

Santiago Acevedo, Alessandro Laio, Marco Baroni

机构 * Catalan Institute of Research and Advanced Studies (ICREA) and Universitat Pompeu Fabra (UPF)(加泰罗尼亚研究与高级科学研究所(ICREA)和庞培法华大学(UPF))

专题命中 知识编辑与模型理解 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本研究通过平均共享句法结构或语义的句子隐藏表示向量,发现大型语言模型(以DeepSeek-V3为例)的内部层表示中句法和语义信息至少部分线性编码,且两者编码轮廓不同,可一定程度解耦。

Comments Published as conference paper at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24946 2026-05-26 cs.CV 83%

Interpretability Transfer from Language to Vision via Sparse Autoencoders

通过稀疏自编码器实现从语言到视觉的可解释性迁移

Alexey Kravets, Da Li, Chuan Li, Da Chen, Vinay P. Namboodiri

机构 * University of Bath, UK(巴斯大学) Lambda, Inc.(Lambda公司) Samsung AI Centre Cambridge(三星AI研究中心)

专题命中 知识编辑与模型理解 :LLM(summary_cn,abstract);language model(abstract)

AI总结 提出VISTA框架,通过约束视觉投影器将视觉token映射到LLM的文本SAE空间,实现无需专用视觉SAE的视觉可解释性,并在对象移除和替换任务上分别提升35%和47%。

Journal ref ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23765 2026-05-08 cs.CL cs.AI cs.LG 83%

Knowledge-Level Consistency Reinforcement Learning: Dual-Fact Alignment for Long-Form Factuality

知识级一致性强化学习:长形式事实性的双事实对齐

Junliang Li, Yucheng Wang, Yan Chen, Yu Ran, Ruiqing Zhang, Jing Liu, Hua Wu, Haifeng Wang

机构 * Baidu Inc.(百度公司)

专题命中 知识编辑与模型理解 :RLHF(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出KLCF框架,通过双事实对齐机制提升长形式生成的事实性,有效缓解幻觉和保守倾向,提升精度与召回率。

Comments 32 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22779 2026-04-28 cs.LG cs.AI cs.CL 83%

KARL: Mitigating Hallucinations in LLMs via Knowledge-Boundary-Aware Reinforcement Learning

KARL:通过知识边界感知强化学习缓解大语言模型的幻觉

Cheng Gao, Cheng Huang, Kangyang Luo, Ziqing Qiao, Shuzheng Si, Huimin Chen, Chaojun Xiao, Maosong Sun

机构 * Tsinghua University(清华大学)

专题命中 知识编辑与模型理解 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 KARL通过知识边界感知强化学习框架,动态调整大语言模型的回避行为,有效抑制幻觉并保持高精度。

Comments 21 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06032 2026-01-13 cs.HC 83%

Applied Theory of Mind and Large Language Models -- how good is ChatGPT at solving social vignettes?

应用认知理论与大语言模型 -- ChatGPT在解决社会情境题方面表现如何?

Anna Katharina Holl-Etten, Nina Schnaderbeck, Elizaveta Kosareva, Leonhard Aron Prattke, Ralph Krueger, Lisa Marie Warner, Nora C. Vetter

专题命中 知识编辑与模型理解 :large language model(title);language model(title)

AI总结 研究评估了GPT-4在解决社会情境题方面的表现,发现其在高阶认知理论任务中接近人类水平,但在不确定性标记使用上仍需进一步优化。

Comments 40 pages, 6 figures, 3 supplements

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00826 2025-08-12 cs.CL cs.AI cs.LG 83%

HERGC: Heterogeneous Experts Representation and Generative Completion for Multimodal Knowledge Graphs

Yongkang Xiao, Rui Zhang

机构 * University of Minnesota(明尼苏达大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.18696 2023-10-31 cs.CL cs.AI cs.LG 83%

Probing LLMs for Joint Encoding of Linguistic Categories

Giulio Starace, Konstantinos Papakostas, Rochelle Choenni, Apostolos Panagiotopoulos, Matteo Rosati, Alina Leidinger, Ekaterina Shutova

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);pretraining(abstract)

Comments Accepted in EMNLP Findings 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.21014 2026-07-22 cs.LG cs.CL 版本更新 83%

CLT-Forge: A Scalable Library for Cross-Layer Transcoders and Attribution Graphs

CLT-Forge:一种可扩展的跨层转码器和归因图库

Florent Draye, Vedant Palit, Abir Harrasse, Tung-Yu Wu, Jiarui Liu, Punya Syon Pandey, Roderick Wu, Chih-Hao Hsu, Terry Jingchen Zhang, Zhijing Jin, Bernhard Schölkopf

机构 * Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) Jinesis AI Lab(Jinesis人工智能实验室) University of Toronto(多伦多大学) Vector Institute(向量研究所) CMU(卡内基梅隆大学) EuroSafeAI ELLIS Institute Tübingen(图宾根ELLIS研究所)

专题命中 知识编辑与模型理解 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出CLT-Forge库,通过分布式训练、模型分片和压缩激活缓存提升跨层转码器的可扩展性,提供统一的可解释性流程和可视化接口,解决归因图的冗余问题。

Comments 9 pages, 7 figures, 1 table. Code: https://github.com/LLM-Interp/CLT-Forge. Demonstration video: https://youtu.be/6ptrrLawTl8

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.17848 2026-08-19 cs.LG cs.SI 新提交 83%

MoRAX: Mobility-based Representation Augmentation for Geospatial Foundation Models

MoRAX:面向地理空间基础模型的基于移动性的表示增强方法

Ya Wen, Jixuan Cai, Yulun Zhou, Alec Kirkley

机构 * The University of Hong Kong(香港大学) The Chinese University of Hong Kong(香港中文大学)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.LG

AI总结 MoRAX是增强地理空间基础模型的轻量级框架,利用人类移动性数据补充区域功能结构,其教师模型在多国多城市的八项预测任务中优于基线,学生模型性能接近教师,可实现零样本部署与跨国家迁移。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.06394 2026-08-10 cs.AI 新提交 83%

Towards Multi-Label Graph Foundation Models: from Single-Vector Representation Learning to Multi-Semantic Basis Learning

面向多标签图基础模型:从单向量表示学习到多语义基学习

Dongxiao He, Jiayu Zhang, Jitao Zhao, Yi Wang, Di Jin

机构 * Tianjin University(天津大学)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI

AI总结 针对现有图基础模型单标签假设导致多标签节点语义纠缠的问题,提出MSB-GFM框架,通过多语义基表示学习与语义-结构双通道架构实现跨域多标签节点分类,经实验验证有效。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.16868 2026-08-06 cs.AI 版本更新 83%

Beyond Semantic Equivalence: Logical Graphs for LLM Uncertainty Quantification

超越语义等价:用于大语言模型不确定性量化的逻辑图

Yanni Dong, Minghua Liu, Meilin Zhu, Xiaowei Huang, Lijun Zhang

机构 * Institute of Intelligent Software(智能软件研究所) Institute of Software, CAS(中国科学院软件研究所) University of Liverpool(利物浦大学) Guangzhou Jiayi Software Technology Co., Ltd.(广州嘉意软件科技有限公司)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究大语言模型输出不可靠问题,提出逻辑图不确定性(LGU)框架,该框架能明确建模答案间逻辑关系,通过聚合概率质量、计算熵等方式改进不确定性估计,在多个问答基准测试中优于现有方法。

Comments 21 pages, 3 figures, 11 tables. Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.03035 2026-08-05 cs.CL 新提交 83%

Language Models Encode the Contextual Truth of Propositions

语言模型编码命题的语境真值

Rupak Sarkar, Pritika Ramu, Rachel Rudinger

专题命中 知识编辑与模型理解 :language model(title);LLM(abstract,abstract_cn);分类 cs.CL

AI总结 该研究探究LLMs对语境真值的编码,发现其会维持语境真值的线性表征,伙伴断言可改变命题真值表征,还区分出两种谄媚形式并量化了其出现频率差异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26845 2026-07-30 cs.LG 新提交 83%

Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models

不确定性下的思考:语言模型中的证据使用与信息寻求

Hua-Dong Xiong, Xinyuan Yan, Ji-An Li, Jingming Xue, Marcelo G. Mattar, Robert C. Wilson

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.LG

AI总结 该研究通过双臂老虎机试验探究大型语言模型在不确定性下的思考机制,发现思考可增强其基于当前证据的行动,但未产生更具信息寻求性的探索策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.25877 2026-07-29 cs.AI 新提交 83%

Runtime Uncertainty Monitoring for LLM-Based Multi-Agent Systems Using Bayesian Networks

基于贝叶斯网络的基于大语言模型的多智能体系统的运行时不确定性监测

Bart Custers, Koorosh Aslansefat

机构 * University of Hull(赫尔大学)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究基于大语言模型的多智能体系统在精算风险建模中的不确定性量化,提出含中央枢纽的多智能体框架,用令牌级对数概率和贝叶斯网络进行不确定性传播,再现基线性能并提供工作流稳定性等见解。

Comments This paper got accepted for Ninth International Workshop on Artificial Intelligence Safety Engineering (WAISE 2026): https://www.waise.org/

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.25244 2026-07-29 cs.AI 新提交 83%

CADENCE: A Cardiac Atom Dictionary for Interpretable Neural Concept Extraction from ECG Foundation Models

CADENCE:用于从心电图基础模型中提取可解释神经概念的心脏原子字典

Yixuan Duan, Arjun Naik, Sadeer Al-Kindi, Wei Qiu

机构 * Rice University(莱斯大学) Houston Methodist Hospital(休斯顿卫理公会医院)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究针对12导联心电图基础模型生理知识不透明问题,提出CADENCE框架,用BatchTopK稀疏自动编码器分解模型,得到稀疏心脏原子,提升临床表型和形态学预测AUROC,能恢复生理关系、验证原子描述,为审查模型知识提供可扩展框架。

Comments 21 pages, 5 main figures, 15 appendix figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07539 2026-07-28 q-bio.NC cs.CL 83%

Training-Driven Representational Geometry Modularization Predicts Brain Alignment in Language Models

训练驱动的表征几何模块化预测语言模型中的脑对齐

Yixuan Liu, Zhiyuan Ma, Likai Tang, Runmin Gan, Xinche Zhang, Jinhao Li, Chao Xie, Sen Song

机构 * School of Biomedical Engineering, Tsinghua University, Beijing, China(生物医学工程学院,清华大学,北京,中国) Tsinghua Laboratory of Brain and Intelligence, Tsinghua University, Beijing, China(脑与智能实验室,清华大学,北京,中国) School of Basic Medical Sciences, Tsinghua University, Beijing, China(基础医学学院,清华大学,北京,中国) Department of Psychological and Cognitive Sciences, Tsinghua University, Beijing, China(心理学与认知科学系,清华大学,北京,中国)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

AI总结 本研究通过训练驱动的表征几何模块化,发现语言模型中的低复杂度模块能更准确预测人类大脑语言网络活动,揭示了表征平滑对神经样语言处理的作用。

Journal ref Proceedings of the Annual Meeting of the Cognitive Science Society, 48 (2026), 167-174

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03190 2026-07-27 cs.CL 版本更新 83%

Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

最大化关键区域的局部熵:基于前缀的局部化大语言模型去敏感化

Naixin Zhai, Pengyang Shao, Binbin Zheng, Yonghui Yang, Fei Shen, Long Bai, Xun Yang

机构 * University of Science and Technology of China(中国科学技术大学) National University of Singapore(新加坡国立大学)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出PALU框架,通过局部熵最大化减少冗余优化,提升大语言模型的去敏感化效果与通用性能。

Comments Accepted to ACL 2026 main

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20778 2026-07-24 cs.LG physics.ao-ph 新提交 83%

Toward Mechanistic Interpretability of an AI Foundation Model Fine-Tuned for Atmospheric Chemistry

迈向针对大气化学进行微调的人工智能基础模型的机理可解释性

Jason Y. Hu, Ivan Higuera-Mendieta, Patrick Obin Sturm, Makoto M. Kelp

机构 * Stanford University(斯坦福大学) University of Southern California(南加州大学) University of Utah(犹他大学)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);pretraining(abstract);分类 cs.LG

AI总结 研究针对大气化学微调的人工智能基础模型学到了什么,通过对微软Aurora模型施加化学扰动、检查内部表示等方法,发现其虽捕捉到部分臭氧响应但未执行化学约束,提供了测试模型是否学大气化学的框架,强调预测应依内部机制评判。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19128 2026-07-22 cs.LG 新提交 83%

One Model, Many Graphs: Learning over Attributed Graphs across Heterogeneous Modalities with Vision-Language Models

一个模型,多种图:使用视觉语言模型在异构模态的属性图上进行学习

Jiayi Yang, Yifang Chen, Yuanfu Sun, Jiajin Liu, Qiaoyu Tan

机构 * New York University Shanghai(纽约大学上海分校) New York University(纽约大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 研究利用视觉语言模型解决属性图模态异质性问题,提出OMG-VLM框架,以预训练VLM为共享主干,引入结构感知图适配器,在节点分类和链接预测等任务中表现出色,优于现有基线且泛化能力强。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.14197 2026-07-17 cs.AI 新提交 83%

How Artificial Intelligence LLM Engines Shape the Global Conflict Information Environment

人工智能大语言模型引擎如何塑造全球冲突信息环境

Jason Miklian

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究人工智能大语言模型引擎对全球冲突信息环境的塑造,通过向五个引擎询问28场冲突问题并分析答案及相关网站,发现可检索记录稀少会致引擎出错,存在生成引擎优化源头优化,探讨了研究意义、政策影响及未来机遇挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.06223 2026-07-16 cs.AI 版本更新 83%

From Reward-Hack Activations to Agentic Risk States: Context-Calibrated Mechanistic Monitoring in LLM Agents

从奖励黑客激活到智能体风险状态:LLM智能体中的上下文校准机制监控

Patrick Wilhelm, Odej Kao

机构 * University of Cambridge(剑桥大学)

专题命中 知识编辑与模型理解 :LLM(title,title_cn);分类 cs.AI

AI总结 本研究通过分析ReAct风格智能体在Gameable ALFWorld和WebShop环境中的奖励黑客行为,提出结合激活状态、熵和决策上下文的上下文校准监控方法,以更准确评估智能体风险。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.10187 2026-07-14 cs.LG cs.SE 新提交 83%

Knowledge-Conditioned, Single-Pass LLM Synthesis of Executable Unity Game Scenes: A Compiler Error Census across 26 Goal Playable Concepts

知识条件下的可执行Unity游戏场景单遍大语言模型合成:26个目标可玩概念的编译器错误普查

Hugh Xuechen Liu, Kıvanç Tatar

机构 * Chalmers University of Technology(查尔姆斯理工大学) University of Gothenburg(哥德堡大学)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 研究大语言模型为Unity游戏场景编写代码,去除迭代修复循环进行单遍生成评估,通过对多个模型、模式等生成的代码分析编译器错误,发现瓶颈是缺少引擎特定知识,按需求对目标模式排序展示单遍生成断点。

Comments Substantially reframed and extended version of arXiv:2603.07101, with new experiments, figures, and analysis

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.08403 2026-07-10 cs.AI 新提交 83%

Game Theory Driven Multi-Agent Framework Mitigates Language Model Hallucination

博弈论驱动的多智能体框架减轻语言模型幻觉

Runzhe Liu, Biquan Bie, Zihao Wang, Yuchao Ma, Yexin Liu, Xinghai Li, Harry Yang, Wenbo Yang, Jinzhe Cao, Shengyang Tao

机构 * Dalian University of Technology(大连理工大学) The Hong Kong University of Science and Technology(香港科技大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.AI

AI总结 研究针对轻量级大语言模型在科学领域应用受限问题,提出基于博弈论的G-Frame多智能体框架,经结构化推理合成语料库训练出OmniChem模型,性能与GPT 4o mini相当且幻觉大幅减少,为科学领域知识发现提供新途径。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.03657 2026-07-07 cs.CV cs.AI 新提交 83%

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

ViPo-MLLM:用于无注释手语翻译的视觉-姿势多模态大语言模型

Ahmed Abul Hasanaath, Bicheng Xu, Mir Rayat Imtiaz Hossain, Leonid Sigal, Hamzah Luqman

机构 * King Fahd University of Petroleum and Minerals(国王法赫德石油矿物大学) University of British Columbia(不列颠哥伦比亚大学)

专题命中 知识编辑与模型理解 :LLM(title,abstract);language model(abstract);分类 cs.AI

AI总结 研究尝试无注释手语翻译,提出ViPo-MLLM框架结合时空RGB和人体姿势特征,用专用编码器与交叉模态注意力,经结构化提示和训练后由大语言模型处理。该模型在数据集取得新成果,验证了机制有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏