arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7539 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7539 篇

2511.14780 2025-11-20 cs.AI 83%

Ask WhAI:Probing Belief Formation in Role-Primed LLM Agents

Keith Moore, Jun W. Kim, David Lyu, Jeffrey Heo, Ehsan Adeli

机构 * Department of Biomedical Data Science, Stanford University(生物医学数据科学系,斯坦福大学)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

Comments Preprint. Accepted for publication at AIAS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03156 2025-11-14 cs.CL 83%

Neural Correlates of Language Models Are Specific to Human Language

Iñigo Parra

机构 * Department of Linguistics University of California, Berkeley(语言学系加州大学伯克利分校)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments To be presented at NeurIPS 2025 Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09684 2025-11-06 cs.CL 83%

Inv-Entropy: A Fully Probabilistic Framework for Uncertainty Quantification in Language Models

Haoyi Song, Ruihan Ji, Naichen Shi, Fan Lai, Raed Al Kontar

机构 * University of Michigan(密歇根大学) University of Minnesota(明尼苏达大学) Northwestern University(西北大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

Journal ref NeurIPS, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02759 2025-11-05 cs.AI cs.SY eess.SY 83%

LLM-Supported Formal Knowledge Representation for Enhancing Control Engineering Content with an Interactive Semantic Layer

Julius Fiedler, Carsten Knoll, Klaus Röbenack

机构 * Institute of Control Theory, TU Dresden(控制理论研究所,德累斯顿技术大学) Chair of Fundamentals of Electrical Engineering, TU Dresden(电气工程基础教授,德累斯顿技术大学)

专题命中 知识编辑与模型理解 :LLM(title,abstract);language model(abstract);分类 cs.AI

Comments 4 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00620 2025-11-04 cs.CL 83%

Certain but not Probable? Differentiating Certainty from Probability in LLM Token Outputs for Probabilistic Scenarios

Autumn Toney-Wails, Ryan Wails

机构 * SciTech Strategies, Inc.(SciTech Strategies公司) Georgetown University(乔治城大学)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

Comments To appear at the Second Workshop on Uncertainty-Aware NLP @EMNLP 2025 (UncertaiNLP '25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20278 2025-11-03 q-bio.QM cs.LG 83%

The cell as a token: high-dimensional geometry in language models and cell embeddings

William Gilpin

机构 * Department of Physics, The University of Texas at Austin, Austin, Texas 78712, USA(德克萨斯大学奥斯汀分校物理系)

专题命中 知识编辑与模型理解 :language model(title);foundation model(abstract);pretraining(abstract);分类 cs.LG

Comments 4 pages, 2 figures

Journal ref Bioinformatics (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19318 2025-10-23 cs.CL 83%

HAD: HAllucination Detection Language Models Based on a Comprehensive Hallucination Taxonomy

Fan Xu, Xinyu Hu, Zhenghan Yu, Li Lin, Xu Zhang, Yang Zhang, Wei Zhou, Jinjie Gu, Xiaojun Wan

机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学计算机技术研究院) Alibaba Group(阿里巴巴集团) Fudan University(复旦大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17148 2025-10-22 cs.SE cs.AI 83%

LLM Agents for Interactive Exploration of Historical Cadastre Data: Framework and Application to Venice

Tristan Karch, Jakhongir Saydaliev, Isabella Di Lenardo, Frédéric Kaplan

机构 * DH-Lab, EPFL, Lausanne, Switzerland(DH实验室,日内瓦联邦理工学院,洛桑,瑞士)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

Comments Accepted in Cambridge press - Computational Humanities Research 2025

Journal ref Comput. humanit. res. 1 (2025) e11

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15804 2025-10-20 cs.CL 83%

Emergence of Linear Truth Encodings in Language Models

Shauli Ravfogel, Gilad Yehudai, Tal Linzen, Joan Bruna, Alberto Bietti

机构 * New York University(纽约大学) Flatiron Institute(Flatiron研究所)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments Accepted in Neurips 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14825 2025-10-17 cs.LG 83%

Programmatic Representation Learning with Language Models

Gabriel Poesia, Georgia Gabriela Sampaio

机构 * Kempner Institute at Harvard University(哈佛大学凯普纳研究所) Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.LG

Comments Code available at https://github.com/gpoesia/leapr/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12587 2025-10-15 cs.CL 83%

Teaching Language Models to Faithfully Express their Uncertainty

Bryan Eikema, Evgenia Ilia, José G. C. de Souza, Chrysoula Zerva, Wilker Aziz

机构 * University of Amsterdam(阿姆斯特丹大学) Outsystems(Outsystems公司) Instituto de Telecomunicações(电信研究所) Instituto Superior Técnico, Universidade de Lisboa(里斯本大学电信研究所)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10331 2025-10-14 cs.AI 83%

LLM-Friendly Knowledge Representation for Customer Support

Hanchen Su, Wei Luo, Wei Han, Yu Elaine Liu, Yufeng Wayne Zhang, Cen Mia Zhao, Ying Joy Zhang, Yashar Mehdad

机构 * Airbnb Inc.(Airbnb公司)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.12032 2025-10-10 cs.CV cs.AI 83%

FSFM: A Generalizable Face Security Foundation Model via Self-Supervised Facial Representation Learning

Gaojian Wang, Feng Lin, Tong Wu, Zhenguang Liu, Zhongjie Ba, Kui Ren

机构 * State Key Laboratory of Blockchain and Data Security(区块链与数据安全国家重点实验室) Hangzhou High-Tech Zone (Binjiang) Institute of Blockchain and Data Security(杭州高新技术区(滨江区)区块链与数据安全研究院)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);pretraining(abstract);分类 cs.AI

Comments 21 pages, 11 figures, project page: https://fsfm-3c.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10246 2025-10-07 cs.LG 83%

Detecting LLM Hallucination Through Layer-wise Information Deficiency: Analysis of Ambiguous Prompts and Unanswerable Questions

Hazel Kim, Tom A. Lamb, Adel Bibi, Philip Torr, Yarin Gal

机构 * University of Oxford(牛津大学)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted to EMNLP(main)2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12113 2025-10-06 cs.CR cs.AI 83%

Semantic Preprocessing for LLM-based Malware Analysis

Benjamin Marais, Tony Quertier, Grégoire Barrue

机构 * Orange Innovation(Orange创新)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01028 2025-10-02 cs.CL stat.ME 83%

Syntax-Guided Diffusion Language Models with User-Integrated Personalization

Ruqian Zhang, Yijiao Zhang, Juan Shen, Zhongyi Zhu, Annie Qu

机构 * Department of Statistics and Data Science, Fudan University(复旦大学统计与数据科学系) Department of Statistics and Applied Probability, University of California, Santa Barbara(加州大学圣巴巴拉分校统计与应用概率系)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25525 2025-10-01 cs.CR cs.LG 83%

Defeating Cerberus: Concept-Guided Privacy-Leakage Mitigation in Multimodal Language Models

Boyang Zhang, Istemi Ekin Akkus, Ruichuan Chen, Alice Dethise, Klaus Satzke, Ivica Rimac, Yang Zhang

机构 * CISPA Helmholtz Center for Information Security(CISPA赫尔姆霍茨信息安全中心) Nokia Bell Labs(诺基亚贝尔实验室)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05652 2025-09-23 cs.CR cs.CL 83%

Sugar-Coated Poison: Benign Generation Unlocks LLM Jailbreaking

Yu-Hang Wu, Yu-Jie Xiong, Hao Zhang, Jia-Chen Zhang, Zheng Zhou

机构 * Shanghai University of Engineering Science(上海工程技术大学) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted by EMNLP2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13154 2025-09-17 cs.CL 83%

LLM Hallucination Detection: A Fast Fourier Transform Method Based on Hidden Layer Temporal Signals

Jinxin Li, Gang Tu, ShengYu Cheng, Junjie Hu, Jinting Wang, Rui Chen, Zhilong Zhou, Dongbo Shan

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12519 2025-09-17 cs.CE cs.CL q-fin.CP 83%

Context-Aware Language Models for Forecasting Market Impact from Sequences of Financial News

Ross Koval, Nicholas Andrews, Xifeng Yan

机构 * University of California, Santa Barbara(加州大学圣巴巴拉分校) Johns Hopkins University(约翰霍普金斯大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12065 2025-09-16 cs.CL 83%

Steering Language Models in Multi-Token Generation: A Case Study on Tense and Aspect

Alina Klerings, Jannik Brinkmann, Daniel Ruffinelli, Simone Ponzetto

机构 * University of Mannheim(曼海姆大学) Technical University Clausthal(克劳斯泰尔大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments to be published in The 2025 Conference on Empirical Methods in Natural Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04664 2025-09-08 cs.CL 83%

Why Language Models Hallucinate

Adam Tauman Kalai, Ofir Nachum, Santosh S. Vempala, Edwin Zhang

机构 * OpenAI Georgia Tech(佐治亚理工学院)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00723 2025-09-03 cs.AI cs.MM 83%

OmniDPO: A Preference Optimization Framework to Address Omni-Modal Hallucination

Junzhe Chen, Tianshu Zhang, Shiyu Huang, Yuwei Niu, Chao Sun, Rongzhou Zhang, Guanyu Zhou, Lijie Wen, Xuming Hu

机构 * Tsinghua University(清华大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) OpenRL Chongqing University(重庆大学)

专题命中 知识编辑与模型理解 :preference optimization(title);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14918 2025-08-22 cs.CY cs.AI 83%

Disentangling the Drivers of LLM Social Conformity: An Uncertainty-Moderated Dual-Process Mechanism

Huixin Zhong, Yanan Liu, Qi Cao, Shijin Wang, Zijing Ye, Zimu Wang, Shiyao Zhang

机构 * Xi’an Jiaotong Liverpool University(西安交通大学利物浦大学) School of Microelectronics, Shanghai University(上海大学微电子学院)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14654 2025-08-21 cs.AI 83%

Entropy-Constrained Strategy Optimization in Urban Floods: A Multi-Agent Framework with LLM and Knowledge Graph Integration

Peilin Ji, Xiao Xue, Simeng Wang, Wenhao Yan

机构 * College of Intelligence and Computing(智能与计算学院)

专题命中 知识编辑与模型理解 :LLM(title,abstract);prompting(abstract);分类 cs.AI

Comments 17 pages including appendix, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16550 2025-08-19 cs.LG stat.ML 83%

A Free Probabilistic Framework for Analyzing the Transformer-based Language Models

Swagatam Das

机构 * Electronics and Communication Sciences Unit, Indian Statistical Institute, Kolkata, India.(印度统计研究所电子与通信科学单元,加尔各答,印度)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16572 2025-07-23 cs.CL 83%

Pixels to Principles: Probing Intuitive Physics Understanding in Multimodal Language Models

Mohamad Ballout, Serwan Jassim, Elia Bruni

机构 * Institute of Cognitive Science, University of Osnabrück(认知科学研究所,奥斯纳布吕克大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10281 2025-07-16 cs.CL 83%

One world, one opinion? The superstar effect in LLM responses

Sofie Goethals, Lauren Rhue

机构 * University of Antwerp(安特卫普大学) Robert H. Smith School of Business(罗伯特·H·史密斯商学院)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

Journal ref https://aclanthology.org/2025.c3nlp-1.pdf#page=100

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10360 2025-07-01 cs.CL cs.IR 83%

Parenting: Optimizing Knowledge Selection of Retrieval-Augmented Language Models with Parameter Decoupling and Tailored Tuning

Yongxin Xu, Ruizhe Zhang, Xinke Jiang, Yujie Feng, Yuzhen Xiao, Xinyu Ma, Runchuan Zhu, Xu Chu, Junfeng Zhao, Yasha Wang

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments Accepted to ACL 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01584 2025-06-30 cs.AI 83%

SENSEI: Semantic Exploration Guided by Foundation Models to Learn Versatile World Models

Cansu Sancaktar, Christian Gumbsch, Andrii Zadaianchuk, Pavel Kolev, Georg Martius

机构 * Autonomous Learning, University of Tübingen(图宾根大学自主学习研究所) Empirical Inference, Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) Neuro-Cognitive Modeling, University of Tübingen(图宾根大学神经认知建模研究所) VISLab, University of Amsterdam(阿姆斯特丹大学视觉实验室)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);language model(abstract);分类 cs.AI

Comments ICML 2025 camera-ready version. Project webpage at https://sites.google.com/view/sensei-paper

详情

展开后加载摘要…

URL PDF HTML 收藏