arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-04-29 至 2026-04-29 共收录 226 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 13 篇

2604.25264 2026-04-29 cs.CR cs.SE 86%

MARD: A Multi-Agent Framework for Robust Android Malware Detection

MARD:一种用于鲁棒Android恶意软件检测的多智能体框架

Xueying Zeng, Youquan Xian, Sihao Liu, Xudong Mou, Yanze Li, Lei Cui, Bo Li

专题命中 知识编辑与模型理解 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract)

AI总结 MARD通过结合LLM的语义理解和传统静态分析,提出多智能体框架,有效降低APK深度分析成本,实现高可解释性检测,F1得分达93.46%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23719 2026-04-29 cs.CL cs.AI 84%

AIPsy-Affect: A Keyword-Free Clinical Stimulus Battery for Mechanistic Interpretability of Emotion in Language Models

AIPsy-Affect:一种无关键词的临床刺激电池,用于语言模型中情绪的机制可解释性

Michael Keeman

机构 * Keido Labs(Keido实验室)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出AIPsy-Affect,一种480项的临床刺激电池,通过叙事情境引发Plutchik的八种基本情绪,消除关键词混淆,支持情绪机制的可解释性研究。

Comments Dataset paper. 12 pages + appendix, 2 figures. Dataset available at https://huggingface.co/datasets/keidolabs/aipsy-affect. MIT license

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12553 2026-04-29 cs.CL cs.AI 81%

Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility

这仅仅是幻想吗?语言模型表示反映了人类对事件可能性的判断

Michael A. Lepori, Jennifer Hu, Ishita Dasgupta, Roma Patel, Thomas Serre, Ellie Pavlick

机构 * Department of Computer Science(计算机科学系) Department of Cognitive Science(认知科学系) Google(谷歌) DeepMind(深度Mind) Brown University(布朗大学) Johns Hopkins University(约翰霍普金斯大学) Department of Cognitive(认知系) Department of Computer Science & Psychological Sciences(计算机科学与心理学科学系)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文通过分析语言模型的模态差异向量,揭示其在事件可能性判断上的能力,发现模型随着训练步骤和参数量增加,能更准确地区分模态类别,并与人类判断行为相关联。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25642 2026-04-29 cs.CV cs.AI 79%

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models

预填充阶段干预用于缓解大视觉-语言模型中的幻觉

Chengsheng Zhang, Chenghao Sun, Xinyan Jiang, Wei Li, Xinmei Tian

机构 * University of Science and Technology of China(中国科学技术大学) Shanghai Advanced Research Institute, Chinese Academy of Sciences(上海先进研究院,中国科学院) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

AI总结 本文提出PTI方法,在预填充阶段干预KV缓存以减少幻觉,通过模态感知方向修正错误表示,提升模型可靠性。

Comments Accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.07802 2026-04-29 cs.CV cs.AI 79%

Latent Anomaly Knowledge Excavation: Unveiling Sparse Sensitive Neurons in Vision-Language Models

潜在异常知识挖掘:揭示视觉-语言模型中的稀疏敏感神经元

Shaotian Li, Shangze Li, Chuancheng Shi, Wenhua Wu, Yanqiu Wu, Xiaohan Yu, Fei Shen, Tat-Seng Chua

机构 * Macquarie University(麦考瑞大学) Nanjing University of Science and Technology(南京理工大学) The University of Sydney(悉尼大学) National University of Singapore(新加坡国立大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

AI总结 本文提出LAKE框架,通过挖掘视觉-语言模型中稀疏敏感神经元,实现异常检测的内在可解释性,实验表明其在工业基准上表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21517 2026-04-29 cs.CL cs.AI 62%

Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation

声音、偏见与指称:语音翻译中性别可解释性研究

Lina Conti, Dennis Fucci, Marco Gaido, Matteo Negri, Guillaume Wisniewski, Luisa Bentivogli

机构 * Fondazione Bruno Kessler(布鲁诺·科塞拉基金会) University of Trento(特伦托大学) Laboratoire de Linguistique Formelle, Université Paris Cité, CNRS(巴黎城市大学语言学实验室,法国国家科学研究中心)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

AI总结 研究探讨语音翻译模型如何根据语音特征分配指称词性别,发现模型通过第一人称代词链接性别化术语与说话者,利用频谱分布而非音高集中信息进行性别判断。

Comments Accepted to LREC 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25902 2026-04-29 cs.CL cs.AI cs.LG 56%

Toward a Functional Geometric Algebra for Natural Language Semantics

迈向自然语言语义的函数几何代数

James Pustejovsky

机构 * Computer Science Department(计算机科学系)

专题命中 知识编辑与模型理解 :分类 cs.CL、cs.AI、cs.LG;language model(comments)

AI总结 本文提出函数几何代数框架,利用克莱因代数提升语义表示的结构组织性,解决传统线性代数在组合语义、类型敏感性和可解释性上的局限。

Comments 43 pages. Keywords: geometric algebra, Clifford algebra, compositional semantics, natural language semantics, type coercion, multivector representations, graded type system, Generative Lexicon, neural language models, distributional semantics

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25102 2026-04-29 cs.CV 50%

One Perturbation, Two Failure Modes: Probing VLM Safety via Embedding-Guided Typographic Perturbations

一个扰动,两种失效模式:通过嵌入引导的字形扰动探测VLM安全性

Ravikumar Balakrishnan, Sanket Mendapara

机构 * Cisco Systems(思科系统)

专题命中 知识编辑与模型理解 :language model(abstract)

AI总结 本文通过实证研究揭示多模态嵌入距离对VLM攻击成功率的预测作用,并提出基于嵌入引导的字形扰动方法,验证了可读性与安全对齐的交互影响。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 其他LLM 8 篇

2509.13400 2026-04-29 cs.CY cs.AI 92%

Justice in Judgment: Unveiling (Hidden) Bias in LLM-assisted Peer Reviews

审判中的公正:揭示(隐藏)在LLM辅助同行评审中的偏见

Sai Suresh Macharla Vasu, Ivaxi Sheth, Hui-Po Wang, Ruta Binkyte, Mario Fritz

机构 * CISPA Helmholtz Center for Information Security(CISPA 欧洲信息安全研究中心) Saarland University(萨尔兰大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究LLM生成的同行评审中的偏见,通过操控作者元数据发现机构偏见、资历偏好和性别影响,揭示隐性偏见在软评分中的显现。

Comments Findings of ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06778 2026-04-29 cs.CL cs.AI 88%

Large Language Models Are Effective Human Annotation Assistants, But Not Good Independent Annotators

大语言模型是有效的标注助手,但不是好的独立标注者

Feng Gu, Zongxia Li, Carlos Rafael Colon, Benjamin Evans, Ishani Mondal, Jordan Lee Boyd-Graber

机构 * Department of Computer Science, University of Maryland(大学计算机科学系) National Consortium for the Study of Terrorism and Responses to Terrorism(反恐与反恐响应国家研究中心)

专题命中 其他LLM :large language model(title);language model(title);LLM(abstract,abstract_cn);分类 cs.CL、cs.AI

AI总结 本文研究了大语言模型在事件标注中的有效性,发现其在辅助专家标注时表现优于传统方法,但独立标注仍不理想。

Comments 9 pages, 4 figures

Journal ref ACL 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06587 2026-04-29 cs.CL cs.AI cs.HC 88%

TouchAI: Exploring human-AI perceptual alignment in touch through language model representations

TouchAI: 探索通过语言模型表示的触觉人机感知对齐

Shu Zhong, Elia Gatti, Youngjun Cho, Marianna Obrist

机构 * Department of Computer Science, University College London(伦敦大学学院计算机科学系)

专题命中 其他LLM :language model(title,abstract);LLM(abstract,abstract_cn);large language model(abstract);分类 cs.CL、cs.AI

AI总结 本文通过'textile hand'任务研究大型语言模型在触觉感知对齐中的表现,发现不同织物样本的对齐程度差异显著,揭示了人机感知对齐的挑战与潜力。

Comments Accepted at IJHCS

Journal ref International Journal of Human-Computer Studies 210 (2026) 103765

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25601 2026-04-29 cs.HC cs.AI 84%

Emotive Architectures: The Role of LLMs in Adjusting Work Environments

情感建筑:大型语言模型在调整工作环境中的作用

Lara Vartziotis, Tina Vartziotis, Frank Beutenmueller, Stella Salta, Konstantinos Moraitis, Miltiadis Katsaros, Sotirios Kotsopoulos

机构 * National Technical University of Athens(希腊国家技术大学) TWT GmbH Science & Innovation(TWT GmbH 科技与创新部门) Massachusetts Institute of Technology(麻省理工学院)

专题命中 其他LLM :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文探讨了LLM在融合物理与虚拟环境中的应用,通过实时调整光照、声学等参数,提升用户专注度与幸福感,并提出伦理考量与包容性设计的重要性。

Comments 19 pages, 1 Table

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25846 2026-04-29 cs.CR cs.AI 81%

Towards Agentic Investigation of Security Alerts

朝向安全警报的代理调查

Even Eilertsen, Vasileios Mavroeidis, Gudmund Grov

机构 * University of Oslo(奥斯陆大学) Norwegian Defence Research Establishment (FFI)(挪威国防研究机构(FFI))

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种利用大语言模型自动化安全警报初步调查的代理工作流,通过预定义查询和结构化工具访问提高准确性,减少人工工作量。

Comments 10 pages, 3 figures, 4 tables. Accepted at the 2025 IEEE International Conference on Big Data (BigData)

Journal ref Proc. 2025 IEEE Int. Conf. on Big Data (BigData), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.07847 2026-04-29 cs.CL cs.AI 73%

From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems

从歧义到准确性:核心指代消解对检索增强生成系统的影响

Youngjoon Jang, Seongtae Hong, Junyoung Son, Sungjin Park, Chanjun Park, Heuiseok Lim

机构 * Korea University(韩国大学) Naver Corp(Naver公司)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 研究探讨了核心指代消解对检索增强生成系统中文档检索和生成性能的影响,发现消解能提升检索效果和问答性能,尤其对小模型有显著帮助。

Comments ACL 2025 SRW

Journal ref https://aclanthology.org/2025.acl-srw.27

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25599 2026-04-29 cs.SE cs.LG 57%

PLMGH: What Matters in PLM-GNN Hybrids for Code Classification and Vulnerability Detection

PLMGH:PLM-GNN混合在代码分类和漏洞检测中的关键因素

Mohamed Taoufik Kaouthar El Idrissi, Edward Zulkoski, Mohammad Hamdaqa

机构 * Polytechnique Montréal(蒙特利尔理工学院) Quantstamp

专题命中 其他LLM :language model(abstract);分类 cs.LG

AI总结 本文通过系统配对三种代码专用PLM与三种基础GNN架构,研究PLM-GNN混合在代码分类和漏洞检测中的性能,发现混合模型优于GNN-only基线,且PLM选择比GNN架构影响更大。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11567 2026-04-29 cs.CY 50%

Consumer Law for AI Agents

人工智能代理的消费者法

Christoph Busch

专题命中 其他LLM :language model(abstract)

AI总结 本文探讨欧盟现行消费者法是否能适应人工智能代理带来的变革,分析其对电子商务和人类中心消费法的挑战,并提出未来兼顾人类与机器的消费者法框架。

Journal ref German Law Journal 26 (2025) 1367-1382

详情

展开后加载摘要…

URL PDF HTML 收藏