arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12169 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12169 篇

2605.05443 2026-05-12 cs.CL cs.AI 84%

SLAM: Structural Linguistic Activation Marking for Language Models

SLAM:语言模型的结构语言激活标记

Fabrice Harel-Canada, Amit Sahai

机构 * University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 其他LLM :language model(title);LLM(abstract,abstract_cn);分类 cs.CL、cs.AI

AI总结 SLAM通过在结构几何中嵌入水印而非词频,实现了高检测准确率且低质量损失,优于现有方案。

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.10627 2026-04-14 cs.CL cs.AI cs.CE 84%

Computational Lesions in Multilingual Language Models Separate Shared and Language-specific Brain Alignment

多语言语言模型中的计算病变分离共享和语言特定的脑对齐

Yang Cui, Jingyuan Sun, Yizheng Sun, Yifan Wang, Yunhao Zhang, Jixing Li, Shaonan Wang, Hongpeng Zhou, John Hale, Chengqing Zong, Goran Nenadic

机构 * The University of Manchester(曼彻斯特大学) City University of Hong Kong(香港城市大学) The Hong Kong Polytechnic University(香港理工大学) Institute of Automation, CAS(中国科学院自动化研究所) University of Chinese Academy of Sciences(中国科学院大学) Johns Hopkins University(约翰霍普金斯大学)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 研究通过多语言大语言模型创建计算病变,揭示语言处理的共享与特定机制,发现共享核心减少整体脑编码相关性,而语言特定病变保持跨语言分离但削弱匹配母语的脑预测性。

Comments 23 pages, 5 figures, Journal format

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.18474 2026-04-10 cs.CL cs.AI 84%

WASD: Locating Critical Neurons as Sufficient Conditions for Explaining and Controlling LLM Behavior

WASD:定位关键神经元作为解释和控制大语言模型行为的充分条件

Haonan Yu, Junhao Liu, Zhenyu Yan, Haoran Lin, Xin Zhang

机构 * Key Lab of High Confidence Software Technologies (Peking University), Ministry of Education(高可信软件技术教育部重点实验室(北京大学)) School of Computer Science, Peking University(北京大学计算机学院)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 WASD通过识别生成token的神经条件,提供更稳定准确的解释,并通过案例验证了控制模型行为的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07139 2026-03-19 cs.CL cs.LG 84%

Byte-token Enhanced Language Models for Temporal Point Processes Analysis

增强字节令牌的语言模型用于时间点过程分析

Quyu Kong, Yixuan Zhang, Yang Liu, Panrong Tong, Enqi Liu, Feng Zhou

机构 * Independent Researcher(独立研究者) Southeast University(东南大学) Center for Applied Statistics and School of Statistics, Renmin University of China(应用统计中心和中国人民大学统计学院)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出Language-TPP框架,结合时间点过程与大语言模型,通过新颖的时序编码机制将连续时间区间转换为字节令牌,提升Web事件序列建模性能,实现事件时间与类型预测的最优表现。

Comments WWW 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19422 2026-03-16 cs.LG cs.CL 84%

LLM Unlearning with LLM Beliefs

基于LLM信念的去学习

Kemou Li, Qizhou Wang, Yue Wang, Fengpeng Li, Jun Liu, Bo Han, Jiantao Zhou

机构 * State Key Laboratory of Internet of Things for Smart City, University of Macau(物联网智能城市国家重点实验室,澳门大学) TMLR Group, Department of Computer Science, Hong Kong Baptist University(TMLR集团,香港 Baptist大学计算机科学系) Imperfect Information Learning Team, RIKEN Center for Advanced Intelligence Project(不完美信息学习团队,RIKEN高级智能项目中心) PRADA Lab, King Abdullah University of Science and Technology(PRADA实验室,国王阿卜杜勒阿齐兹大学科学与技术学院) National Institute of Informatics(国家信息研究所)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出BS框架,通过结合模型自身高置信度生成(即模型信念)来对抗去学习中的挤压效应,从而更彻底地实现遗忘并保持实用性。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03329 2026-03-05 cs.CL cs.AI 84%

AutoHarness: improving LLM agents by automatically synthesizing a code harness

AutoHarness: 通过自动合成代码框架提升大语言模型代理

Xinghua Lou, Miguel Lázaro-Gredilla, Antoine Dedieu, Carter Wendelken, Wolfgang Lehrach, Kevin P. Murphy

机构 * Google DeepMind(谷歌DeepMind)

专题命中 其他LLM :LLM(title,abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 通过自动合成代码框架,Gemini-2.5-Flash在多个游戏中超越大模型,提升性能并降低成本。

Comments agent harness, code synthesis, self-improvement, code-as-policy, text games

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02983 2026-03-04 cs.CR cs.AI cs.CL 84%

Contextualized Privacy Defense for LLM Agents

上下文化隐私防御用于大语言模型代理

Yule Wen, Yanzhe Zhang, Jianxun Lian, Xiaoyuan Yi, Xing Xie, Diyi Yang

机构 * Tsinghua University(清华大学) Stanford University(斯坦福大学) Microsoft(微软)

专题命中 其他LLM :LLM(title,abstract);prompting(abstract);分类 cs.CL、cs.AI

AI总结 本文提出上下文化防御指导(CDI),通过强化学习优化框架,在大语言模型代理执行中主动塑造隐私保护与有用性之间的平衡。

Comments 25 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00669 2026-03-03 cs.CL cs.AI cs.HC 84%

SSKG Hub: An Expert-Guided Platform for LLM-Empowered Sustainability Standards Knowledge Graphs

SSKG Hub: 一个基于大语言模型的可持续性标准知识图谱专家指导平台

Chaoyue He, Xin Zhou, Xinjia Yu, Lei Zhang, Yan Zhang, Yi Wu, Lei Xiao, Liangyue Li, Di Wang, Hong Xu, Xiaoqiao Wang, Wei Liu, Chunyan Miao

机构 * Alibaba-NTU Global e-Sustainability CorpLab (ANGEL)(阿里-国立大学全球可持续性公司实验室(ANGEL)) Alibaba Group(阿里巴巴集团)

专题命中 其他LLM :LLM(title,abstract);prompting(abstract);分类 cs.CL、cs.AI

AI总结 SSKG Hub通过大语言模型和专家指导构建可持续性标准知识图谱,实现标准到可审计图谱的转化,并提供治理框架和跨图谱融合功能。

Comments 10 pages, 2 figures, 2 tables, submitted to ACL26 System Demo Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25184 2026-02-26 cs.CL cs.AI cs.GT 84%

Incentive-Aligned Multi-Source LLM Summaries

对齐激励的多源大语言模型摘要

Yanchen Jiang, Zhe Feng, Aranyak Mehta

机构 * Harvard University(哈佛大学) Google Research(谷歌研究)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出TTS框架,通过激励对齐提升多源摘要的事实准确性与稳健性,同时保持流畅性。

Comments Accepted at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21752 2026-02-05 cs.CL cs.AI 84%

Semantics as a Shield: Label Disguise Defense (LDD) against Prompt Injection in LLM Sentiment Classification

语义作为盾牌:对抗大语言模型情感分类中提示注入的标签伪装防御(LDD)

Yanxi Li, Ruocheng Shan

机构 * Department of Computer Science, George Washington University(计算机科学系,乔治华盛顿大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出LDD,一种通过语义伪装标签来防御大语言模型情感分类中提示注入攻击的方法,展示了其在不同模型上的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01703 2026-02-03 cs.LG cs.CL 84%

$\textbf{AGT$^{AO}$}$: Robust and Stabilized LLM Unlearning via Adversarial Gating Training with Adaptive Orthogonality

AGT$^{AO}$:通过对抗门控训练与自适应正交性实现鲁棒且稳定的LLM反向学习

Pengyu Li, Lingling Zhang, Zhitao Gao, Yanrui Wu, Yuxuan Dong, Huan Liu, Bifan Wei, Jun Liu

机构 * School of Computer Science and Technology, Xi’an Jiaotong University, China(西安交通大学计算机科学与技术学院) MOE KLINNS Lab, Xi’an Jiaotong University, China(西安交通大学MOE KLINNS实验室) Shaanxi Province Key Laboratory of Big Data Knowledge Engineering, China(陕西省大数据知识工程重点实验室)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 AGT$^{AO}$通过对抗门控训练与自适应正交性,实现LLM反向学习的鲁棒性和实用性平衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16035 2026-01-12 cs.CL cs.AI 84%

Liars' Bench: Evaluating Lie Detectors for Language Models

说谎的检验台:评估语言模型的说谎检测器

Kieron Kretschmar, Walter Laurito, Sharan Maiya, Samuel Marks

机构 * Cadenza Labs(Cadenza实验室) FZI(弗劳恩霍夫研究所) University of Cambridge(剑桥大学)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出LIARS' BENCH,通过多个数据集和模型生成72,863个谎言和诚实响应,评估三种谎言检测技术,揭示现有方法在识别特定类型谎言上的局限性。

Comments *Kieron Kretschmar and Walter Laurito contributed equally to this work. 10 pages, 2 figures; plus appendix. Code at https://github.com/Cadenza-Labs/liars-bench and datasets at https://huggingface.co/datasets/Cadenza-Labs/liars-bench Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08726 2026-01-08 cs.CL cs.AI 84%

Improved LLM Agents for Financial Document Question Answering

改进的金融文档问答大型语言模型代理

Nelvin Tan, Zian Seng, Liang Zhang, Yu-Ching Shih, Dong Yang, Amol Salunkhe

机构 * American Express(美国美国运通)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出改进的金融文档问答代理,通过实验展示其在无 oracle 标签情况下的有效性,并引入更安全的计算代理。

Comments 13 pages, 6 figures. More analysis is added to Appendix C

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.01433 2025-12-29 q-bio.QM cs.CL cs.LG 84%

Enhancing TCR-Peptide Interaction Prediction with Pretrained Language Models and Molecular Representations

利用预训练语言模型和分子表示增强TCR-肽相互作用预测

Cong Qi, Hanzhang Fang, Siqi jiang, Tianxing Hu, Zhi Wei

机构 * New Jersey Institute of Technology(新泽西理工学院)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.LG

AI总结 LANTERN通过结合预训练语言模型和分子表示,提升TCR-肽相互作用预测的准确性和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05162 2025-12-08 stat.ML cs.AI cs.LG math.DS math.PR 84%

How to Tame Your LLM: Semantic Collapse in Continuous Systems

如何驯服你的大语言模型:连续系统中的语义崩溃

C. M. Wyss

机构 * Exolytica AI

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出通过连续状态机理论解释大语言模型中离散符号语义的生成,揭示了语义坍缩与逻辑可解释性的统一。

Comments 35 pages, 1 figure. Exolytica AI Technical Report XTR-2025-01

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.00034 2025-11-18 cs.CL cs.AI 84%

Is Our Chatbot Telling Lies? Assessing Correctness of an LLM-based Dutch Support Chatbot

Herman Lassche, Michiel Overeem, Ayushi Rastogi

机构 * Product Development, AFAS Software(AFAS软件产品开发部) Faculty of Science and Engineering, University of Groningen(格罗宁根大学科学与工程学院)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 10 pages + 2 pages references, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01100 2025-11-05 cs.CL cs.LG 84%

Repetitions are not all alike: distinct mechanisms sustain repetition in language models

Matéo Mahaut, Francesca Franzon

机构 * Universitat Pompeu Fabra(庞培法布拉大学)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08009 2025-10-10 cs.AI cs.LG 84%

Language Models Do Not Embed Numbers Continuously

Alex O. Davies, Roussel Nzoyem, Nirav Ajmeri, Telmo M. Silva Filho

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI、cs.LG

Comments 12 pages, 10 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07613 2025-10-10 cs.CL cs.AI 84%

Vocabulary embeddings organize linguistic structure early in language model training

Isabel Papadimitriou, Jacob Prince

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04498 2025-10-07 cs.CL cs.AI 84%

GenQuest: An LLM-based Text Adventure Game for Language Learners

Qiao Wang, Adnan Labib, Robert Swier, Michael Hofmeyr, Zheng Yuan

机构 * Hosei University(立命馆大学) King’s College London(伦敦大学国王学院) Kindai University(_kindai大学) Tokyo Uni. of Science(东京科学大学) University of Sheffield(谢菲尔德大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Workshop on Wordplay: When Language Meets Games, EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12903 2025-10-07 cs.CL cs.AI 84%

A Stitch in Time Saves Nine: Proactive Self-Refinement for Language Models

Jinyi Han, Xinyi Wang, Haiquan Zhao, Tingyun li, Zishang Jiang, Sihang Jiang, Jiaqing Liang, Xin Lin, Weikang Zhou, Zeye Sun, Fei Yu, Yanghua Xiao

机构 * Shanghai Institute of Artificial Intelligence for Education(上海人工智能教育研究院) East China Normal University(华东师范大学) School of Data Science, Fudan University(复旦大学数据科学学院) College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院) Antgroup(蚂蚁集团)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17348 2025-09-23 cs.CL cs.AI 84%

AIMMerging: Adaptive Iterative Model Merging Using Training Trajectories for Language Model Continual Learning

Yujie Feng, Jian Li, Xiaoyu Dong, Pengfei Xu, Xiaohui Zhou, Yujia Zhang, Zexin LU, Yasha Wang, Alan Zhao, Xu Chu, Xiao-Ming Wu

机构 * Al Technology Center of OVB, Tencent, China(腾讯奥比大学技术中心) The Hong Kong Polytechnic University, Hong Kong S.A.R.(香港理工大学) Peking University, China(北京大学)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.21589 2025-08-15 cs.CL cs.AI cs.DL 84%

DNB-AI-Project at SemEval-2025 Task 5: An LLM-Ensemble Approach for Automated Subject Indexing

Lisa Kluge, Maximilian Kähler

机构 * Deutsche Nationalbibliothek(德国国家图书馆)

专题命中 其他LLM :LLM(title,abstract);prompting(abstract);分类 cs.CL、cs.AI

Comments 11 pages, 4 figures, submitted to SemEval-2025 workshop Task 5: LLMs4Subjects

Journal ref In Proceedings of the 19th International Workshop on Semantic Evaluation (SemEval-2025), pages 1118-1128, Vienna, Austria. Association for Computational Linguistics

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21073 2025-07-16 cs.CL cs.LG 84%

Shared Global and Local Geometry of Language Model Embeddings

Andrew Lee, Melanie Weber, Fernanda Viégas, Martin Wattenberg

机构 * Harvard University(哈佛大学) Google DeepMind(谷歌DeepMind)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10442 2025-07-15 cs.LG cs.AI 84%

Response Wide Shut? Surprising Observations in Basic Vision Language Model Capabilities

Shivam Chandhok, Wan-Cyuan Fan, Vered Shwartz, Vineeth N Balasubramanian, Leonid Sigal

机构 * University of British Columbia(不列颠哥伦比亚大学) Vector Institute for AI(人工智能矢量研究所) IIT Hyderabad(海得拉巴印度理工学院) CIFAR AI Chair(卡尔加里人工智能主席) Microsoft Research India(微软印度研究院)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

Comments Accepted at ACL 2025 (Main Conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06466 2025-07-10 cs.LG cs.AI 84%

Foundation Model Self-Play: Open-Ended Strategy Innovation via Foundation Models

Aaron Dharna, Cong Lu, Jeff Clune

专题命中 其他LLM :foundation model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

Comments 67 pages, accepted to RLC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13305 2025-06-26 cs.CL cs.AI 84%

Computation Mechanism Behind LLM Position Generalization

Chi Han, Heng Ji

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments ACL 2025 Main Long Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15961 2025-06-25 cs.DC cs.AI cs.LG 84%

TrainVerify: Equivalence-Based Verification for Distributed LLM Training

Yunchi Lu, Youshan Miao, Cheng Tan, Peng Huang, Yi Zhu, Xian Zhang, Fan Yang

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00085 2025-06-03 cs.CL cs.AI 84%

COSMIC: Generalized Refusal Direction Identification in LLM Activations

Vincent Siu, Nicholas Crispino, Zihao Yu, Sam Pan, Zhun Wang, Yang Liu, Dawn Song, Chenguang Wang

机构 * Washington University in St. Louis(华盛顿大学圣路易斯分校) University of California, Berkeley(加州大学伯克利分校) University of California, Santa Cruz(加州大学圣克ruz分校)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 9 pages, Accepted to ACL 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11953 2025-05-29 cs.LG cs.AI 84%

Exploring Criteria of Loss Reweighting to Enhance LLM Unlearning

Puning Yang, Qizhou Wang, Zhuo Huang, Tongliang Liu, Chengqi Zhang, Bo Han

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏