arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7539 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7539 篇

2511.04588 2026-04-28 cs.AI cs.CY 84%

Question the Questions: Auditing Representation in Online Deliberative Processes

质疑问题:在线辩论过程中的代表性审计

Soham De, Lodewijk Gelauff, Ashish Goel, Smitha Milli, Ariel Procaccia, Alice Siu

机构 * FAIR at Meta(Meta的FAIR)

专题命中 知识编辑与模型理解 :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出基于公正代表概念的审计框架,用于评估辩论过程中问题集的代表性,通过算法和历史数据验证了LLM在支持辩论过程中的潜力与局限。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23942 2026-04-28 cs.HC cs.AI 84%

What Did They Mean? How LLMs Resolve Ambiguous Social Situations across Perspectives and Roles

他们是什么意思?大语言模型如何在不同视角和角色下解析模糊的社会情境

Qiming Yuan, Linyi Han, Nam Ling, Cihan Ruan

机构 * Santa Clara University(圣克拉拉大学)

专题命中 知识编辑与模型理解 :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究探讨LLM在处理模糊社会情境时的响应方式,发现其倾向于通过叙事一致、规范建议等路径闭合模糊性,且叙述视角影响解读路径,凸显LLM在社会AI设计中保持不确定性的挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22335 2026-04-27 cs.CL 84%

Context-Fidelity Boosting: Enhancing Faithful Generation through Watermark-Inspired Decoding

上下文保真度提升:通过水印启发式解码增强忠实生成

Weixu Zhang, Fanghua Ye, Qiang Gao, Jian Li, Haolun Wu, Yuxing Tian, Sijing Duan, Nan Du, Xiaolong Li, Xue Liu

机构 * Hunyuan AI Digital Human, Tencent(腾讯文深AI数字人) McGill University(麦吉尔大学) Mila - Quebec AI Institute(魁北克人工智能研究所) Wuhan University(武汉大学) University of Montreal(蒙特利尔大学) Tsinghua University(清华大学) MBZUAI

专题命中 知识编辑与模型理解 :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出CFB框架,通过增加源支持标记的生成概率减少大语言模型中的保真度幻觉,采用基于水印技术的logit调整策略,三种增强策略提升生成忠实度,无需重训练即可兼容多种LLM。

Comments Accepted at ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20091 2026-04-17 cs.CL 84%

How Retrieved Context Shapes Internal Representations in RAG

检索上下文如何塑造RAG中的内部表示

Samuel Yeh, Sharon Li

机构 * Department of Computer Science, University of Wisconsin-Madison(威斯康星大学麦迪逊分校计算机科学系)

专题命中 知识编辑与模型理解 :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究通过潜在表示分析检索文档类型对LLM隐藏状态的影响,揭示上下文相关性和分层处理对内部表示的影响,为RAG系统设计提供见解。

Comments ACL 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05727 2026-03-09 cs.CL cs.NA math.NA 84%

Structured Multidimensional Representation Learning for Large Language Models

结构化多维表示学习用于大语言模型

Alaa El Ichi, Khalide Jbilou, Mohamed El Guide, Franck Dufrenois

机构 * Université du Littoral Cote d’Opale(勒阿弗尔海岸大学) LMPA(LMPA研究所) FGSES, University Mohammed VI Polytechnic(穆莱·伊沙克六世理工学院FGSES研究所) LISIC(LISIC研究所)

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.CL

AI总结 本文提出一种基于谱分解的张量变换器,通过结构化多维表示学习压缩大语言模型参数,同时保持模型性能。

Comments 25 pages, 6 figures. Preprint of a journal submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20794 2025-12-25 cs.CL 84%

Investigating Model Editing for Unlearning in Large Language Models

探究大型语言模型中去学习的模型编辑

Shariqah Hossain, Lalana Kagal

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.CL

AI总结 本研究探讨了大型语言模型中去学习的模型编辑方法,通过设计新的编辑目标,展示了模型编辑在去学习任务中的优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27131 2025-11-03 cs.LG 84%

Exploring the Utilities of the Rationales from Large Language Models to Enhance Automated Essay Scoring

Hong Jiao, Hanna Choi, Haowei Hua

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.LG

Comments 12 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19127 2025-10-14 cs.CL 84%

Exploring the Generalizability of Factual Hallucination Mitigation via Enhancing Precise Knowledge Utilization

Siyuan Zhang, Yichi Zhang, Yinpeng Dong, Hang Su

机构 * Dept. of Comp. Sci. and Tech.(计算机科学与技术系) Institute for AI(人工智能研究院) Tsinghua-Bosch Joint ML Center(清华大学-博世联合机器学习中心) THBI Lab(THBI实验室) BNRist Center(BNRist中心) Tsinghua University(清华大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);post-training(abstract)

Comments 33 pages, 17 figures, 21 tables, accepted by EMNLP 2025 as findings paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05664 2025-10-08 cs.AI 84%

Large Language Model-Based Uncertainty-Adjusted Label Extraction for Artificial Intelligence Model Development in Upper Extremity Radiography

Hanna Kreutzer, Anne-Sophie Caselitz, Thomas Dratsch, Daniel Pinto dos Santos, Christiane Kuhl, Daniel Truhn, Sven Nebelung

机构 * Lab for Artificial Intelligence in Medicine, Department of Diagnostic and Interventional Radiology, University Hospital Aachen(人工智能医学实验室,诊断与介入放射科,亚琛大学医院) Department of Diagnostic and Interventional Radiology, University Hospital Aachen(诊断与介入放射科,亚琛大学医院) Institute for Diagnostic and Interventional Radiology, Faculty of Medicine and University Hospital Cologne, University of Cologne(诊断与介入放射学研究所,医学学院和科隆大学医院,科隆大学) Department of Diagnostic and Interventional Radiology, University Medical Center Mainz(诊断与介入放射科,马因茨大学医学中心)

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.AI

Comments 28 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04429 2025-09-22 cs.AI 84%

Activation Space Interventions Can Be Transferred Between Large Language Models

Narmeen Oozeer, Dhruv Nathawani, Nirmalendu Prakash, Michael Lan, Abir Harrasse, Amirali Abdullah

机构 * Nvidia Singapore University of Technology and Design(新加坡技术与设计大学)

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.AI

Comments 75 pages. Accepted to ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13676 2025-09-08 cs.AI 84%

MHSNet:An MoE-based Hierarchical Semantic Representation Network for Accurate Duplicate Resume Detection with Large Language Model

Yu Li, Zulong Chen, Wenjian Xu, Hong Wen, Yipeng Yu, Man Lung Yiu, Yuyu Yin

机构 * Hangzhou Dianzi University(杭州电子科技大学) Alibaba Group(阿里巴巴集团) Zhejiang University of Science(浙江理工大学) Taotian, Alibaba Group(阿里天天,阿里巴巴集团) Department of Computing, Hong Kong Polytechnic University(计算系,香港理工大学)

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05553 2025-08-11 cs.CL 84%

Latent Structure Modulation in Large Language Models Through Stochastic Concept Embedding Transitions

Stefan Whitaker, Colin Sisate, Marcel Windsor, Nikolai Fairweather, Tarquin Goldborough, Oskar Lindenfeld

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.CL

Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24635 2025-06-02 cs.CL 84%

Disentangling Language and Culture for Evaluating Multilingual Large Language Models

Jiahao Ying, Wei Tang, Yiran Zhao, Yixin Cao, Yu Rong, Wenxuan Zhang

机构 * Singapore Management University(新加坡管理学院) University of Science and Technology of China(中国科学技术大学) National University of Singapore(新加坡国立大学) Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究院) DAMO Academy, Alibaba Group(阿里集团大模型实验室) Singapore University of Technology and Design(新加坡科技设计大学)

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.CL

Comments Accepted to ACL 2025 (Main Conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21003 2025-05-28 cs.CL 84%

Uncertainty Unveiled: Can Exposure to More In-context Examples Mitigate Uncertainty for Large Language Models?

Yifei Wang, Yu Sheng, Linjing Li, Daniel Zeng

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院) School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学)

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.CL

Comments Camera-ready versions for ACL 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.12755 2025-04-18 cs.RO cs.AI 84%

Trajectory Adaptation using Large Language Models

Anurag Maurya, Tashmoy Ghosh, Ravi Prakash

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.AI

Comments Accepted to CoRL LangRob workshop 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05395 2025-03-26 cs.CL 84%

Hierarchical Lexical Manifold Projection in Large Language Models: A Novel Mechanism for Multi-Scale Semantic Representation

Natasha Martus, Sebastian Crowther, Maxwell Dorrington, Jonathan Applethwaite, Edgar Tillinghurst, Quentin Birkenshaw, Lukas Petrov, Constance Willoughby

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.CL

Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.02917 2025-03-06 eess.IV cs.AI cs.CV 84%

Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models

Deval Mehta, Yiwen Jiang, Catherine L Jan, Mingguang He, Kshitij Jadhav, Zongyuan Ge

专题命中 知识编辑与模型理解 :language model(title);prompting(title);分类 cs.AI

Comments Accepted to Information Processing in Medical Imaging (IPMI) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12131 2025-02-18 cs.AI 84%

Transformer Dynamics: A neuroscientific approach to interpretability of large language models

Jesseba Fernando, Grigori Guitchounts

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.14603 2024-09-24 cs.AI 84%

Brain Surgery: Ensuring GDPR Compliance in Large Language Models via Concept Erasure

Michele Laurelli

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.06908 2024-07-10 cs.CL cs.CY 84%

Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models

Flor Miriam Plaza-del-Arco, Amanda Cercas Curry, Susanna Paoli, Alba Curry, Dirk Hovy

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.08703 2023-06-22 cs.LG 84%

PAC Prediction Sets for Large Language Models of Code

Adam Khakhar, Stephen Mell, Osbert Bastani

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.LG

Comments Proceedings of the 40th International Conference on Machine Learning

Journal ref PMLR 202, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07209 2026-05-11 cs.CL cs.AI cs.LG 84%

Hallucination Detection via Activations of Open-Weight Proxy Analyzers

通过开放权重代理分析器的激活进行幻觉检测

Akshita Singh, Prabesh Paudel, Siddhartha Roy

机构 * Khoury College of Computer Sciences(科里学院计算机科学学院) Northeastern University(东北大学)

专题命中 知识编辑与模型理解 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出一种代理分析器框架,通过本地部署的小型开放权重模型检测大语言模型的幻觉,采用变压器文本处理特性构建18种特征,并在多个模型架构上训练堆叠集成模型,最终在RAGTruth数据集上取得优于ReDeEP的AUC和F1分数。

Comments 12 pages, 4 figures. Code available at https://github.com/hallu-detect/llm_hallucination_detection

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.08926 2026-08-11 cs.AI cs.LG 新提交 84%

Decoding Phenotypes: A Framework for Fusing Genomic Language Models and Neuroimaging

解码表型:融合基因组语言模型与神经成像的框架

Tianli Tao, Ziyang Wang, Emma Robinson, Rachel Sparks, Le Zhang

机构 * King’s College London(伦敦国王学院) Aston University(阿斯顿大学) University of Birmingham(伯明翰大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI、cs.LG

AI总结 该研究提出多模态框架GeneFuse,整合GLM遗传表征与神经成像特征,在NC vs. MCI、NC vs. AD任务上分别获0.77、0.83的AUROC,性能优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00435 2026-08-03 cs.CL cond-mat.dis-nn cs.AI nlin.CD 版本更新 84%

Escaping Mode Collapse in LLM Generation via Geometric Regulation

通过几何调控逃离大语言模型生成中的模式崩溃

Xin Du, Kumiko Tanaka-Ishii

机构 * Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,地点,国家) School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家) Department of Communications and Computer Engineering, Waseda University, Tokyo, Japan(通信与计算机工程系,早稻田大学,东京,日本) Department of Computer Science and Engineering, Waseda University, Tokyo, Japan(计算机科学与工程系,早稻田大学,东京,日本) Shanghai Research Institute for Intelligent Autonomous Systems, Tongji University, Shanghai, China(智能自主系统上海研究院,同济大学,上海,中国)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文从动力系统视角将模式崩溃解释为几何崩溃,并提出轻量级在线状态空间干预方法RMR(通过低秩阻尼调控Transformer值缓存中的自强化方向),显著降低模式崩溃并实现极低熵率下的稳定生成。

Comments Accepted to ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.27386 2026-07-31 cs.AI cs.LG 新提交 84%

Beyond the Bidirectional Promise: Re-evaluating the Robustness of Diffusion Language Models

超越双向承诺:重新评估扩散语言模型的鲁棒性

Saurabh Yadav, Badri Narayana Patro, Vijay Srinivas Agneeswaran

机构 * Microsoft(微软公司)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI、cs.LG

AI总结 该研究评估了扩散语言模型的鲁棒性,发现其虽能抵御部分对抗攻击但易受自然噪声影响,存在过度自信问题,脆弱性源于解码器路由故障,需将鲁棒性整合入迭代解码循环。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20487 2026-07-27 cs.AI cs.CL 版本更新 84%

Directional Hallucinations: Ideological Drift in News-Grounded LLM Question Answering

定向幻觉:基于新闻的语言模型问答中的意识形态漂移

Chendi Wang, Liam Cunningham, Tom Yishay, Jieying Chen

机构 * Vrije Universiteit Amsterdam(阿姆斯特丹自由大学)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 研究基于新闻的语言模型问答中的意识形态漂移,提出可重复测量框架,利用大量新闻文章对多个模型进行实验,发现幻觉率因模型和话题而异,幻觉内容有向左漂移现象,并探讨了相关意义。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19618 2026-07-23 q-bio.GN cs.AI cs.LG 新提交 84%

Causal dictionary learning reveals and validates transcription-factor binding features in genomic language models

因果字典学习揭示并验证基因组语言模型中的转录因子结合特征

Sarwan Ali

机构 * Columbia University Irving Medical Center(哥伦比亚大学伊文思医疗中心)

专题命中 知识编辑与模型理解 :language model(title,abstract);foundation model(abstract);分类 cs.AI、cs.LG

AI总结 研究针对基因组语言模型内部表示不透明问题,引入结合稀疏字典学习与因果干预的框架,训练自动编码器提取转录因子结合特征,开发去混淆协议并因果验证,为基因组深度学习可解释性提供计算标准。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.09749 2026-07-14 eess.SP cs.AI cs.LG 新提交 84%

MorphologyFM: A Foundation Model for Morphology-Aware Representation Learning from ECG and Pulse Oximetry Waveforms

MorphologyFM:一种用于从心电图和脉搏血氧波形中进行形态感知表示学习的基础模型

Saiyang Feng, Yuanyun Zhang, Shi Li

机构 * University of the Chinese Academy of Sciences(中国科学院大学) Columbia University(哥伦比亚大学)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);pretraining(abstract);分类 cs.AI、cs.LG

AI总结 研究针对现有生理波形方法未保留临床意义波形形态的问题,提出多模态基础模型MorphologyFM,通过形态感知自监督学习目标预训练,结合多种技术学习相关表示,在多下游任务中表现优于其他方法,证明联合建模更具可转移性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.08017 2026-07-10 cs.CL cs.AI 新提交 84%

Can We Trust LLM's Logic? Quantifying Uncertainty, Coherence, and Robustness via a Graph-Based Framework

我们能信任大语言模型的逻辑吗?通过基于图的框架量化不确定性、连贯性和鲁棒性

Riccardo Revalor, Jalees Rehman, Debjit Pal

机构 * Department of Electrical and Computer Engineering(电气与计算机工程系) Department of Biochemistry and Molecular Genetics(生物化学与分子遗传学系) University of Illinois Chicago(伊利诺伊大学芝加哥分校)

专题命中 知识编辑与模型理解 :LLM(title,abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 研究大语言模型推理问题,引入基于图的GRAPHEVAL框架,提出GRCS度量和GSC解码策略,量化不确定性等,发现GRCS与推理忠实度负相关,证明GSC所选路径对推理的重要性。

Comments 42 pages, 14 figures, 12 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.04525 2026-07-07 cs.CL cs.AI 新提交 84%

Language Models Represent and Transform Concepts with Shared Geometry

语言模型通过共享几何表示和转换概念

Zhimin Hu, Lanhao Niu, Sashank Varma

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 研究神经网络中概念表示问题,将概念表示形式化为点云流形、上下文转换为向量场,在大语言模型中实例化该框架,发现模型共享概念表示及上下文转换的通用几何结构。

详情

展开后加载摘要…

URL PDF HTML 收藏