arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-05-04 至 2026-05-04 共收录 14 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 14 篇

2605.00012 2026-05-04 cs.IR cs.AI cs.CL 92%

Exploring LLM biases to manipulate AI search overview

探索LLM偏见以操控AI搜索概述

Roman Smirnov

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);small language model(abstract)

AI总结 研究探索LLM偏见在AI搜索概述系统中的影响,通过强化学习优化搜索片段内容以操控结果,并发现偏见驱动相对优势而非绝对优势。

Comments 14 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18315 2026-05-04 cs.SE cs.AI 90%

Effective LLM Code Refinement via Property-Oriented and Structurally Minimal Feedback

通过属性导向和结构最小反馈的有效LLM代码细化

Lehan He, Zeren Chen, Zhe Zhang, Xiang Gao, Lu Sheng

机构 * School of Software, Beihang University, Beijing, China(北京航空航天大学软件学院) Shanghai AI Laboratory, Shanghai, China(上海人工智能实验室) Shanghai Innovation Institute, Shanghai, China(上海创新研究院)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.AI

AI总结 本文提出属性生成求解器(PGS),通过属性导向和结构最小反馈提升LLM代码生成的正确性,实验显示PGS在pass@1和修复率上均优于其他TDD方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00087 2026-05-04 cs.NI cs.AI cs.CY cs.IR cs.LG 88%

DeGenTWeb: A First Look at LLM-dominant Websites

DeGenTWeb:对以大语言模型为主导的网站的首次观察

Sichang Steven He, Calvin Ardi, Ramesh Govindan, Harsha V. Madhyastha

机构 * University of Southern California(南加州大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文通过DeGenTWeb系统识别出大量由大语言模型生成内容的网站,发现其在Web中的普及率持续增长,但准确识别此类网站面临挑战。

Comments 6 pages, 6 figures, 13 page total; in submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00364 2026-05-04 cs.CL 83%

Unlearning What Matters: Token-Level Attribution for Precise Language Model Unlearning

撤销重要信息:针对语言模型精确撤销的标记级归因

Jiawei Wu, DouDou Zhou

机构 * Department of Statistics and Data Science, National University of Singapore(统计与数据科学系,新加坡国立大学)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

AI总结 本文提出TokenUnlearn框架,通过标记级归因实现精确撤销,结合知识意识信号和熵意识信号生成重要性评分,改进梯度信噪比,实验显示在遗忘效果和效用保持方面优于序列级基线。

Comments 17 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00782 2026-05-04 cs.SE cs.AI 81%

GeoContra: From Fluent GIS Code to Verifiable Spatial Analysis with Geography-Grounded Repair

GeoContra:从流畅的GIS代码到可验证的空间分析与地理基础修复

Yinhao Xiao, Rongbo Xiao, Yihan Zhang

机构 * School of Big Data and Artificial Intelligence, Guangdong University of Finance and Economics(大数据与人工智能学院,广东财经大学) School of Geography and Environment Economics, Guangdong University of Finance & Economics(地理与环境经济学院,广东财经大学) Guangdong Engineering Research Center of Low-Altitude Remote Sensing Intelligent Monitoring(低空遥感智能监测工程研究中心)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI

AI总结 GeoContra通过验证和修复框架提升LLM驱动的GIS工作流空间正确性,通过静态检查、运行时验证和语义验证提高空间分析准确性,实验显示在多个模型上空间正确性显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06572 2026-05-04 cs.CR 80%

Parasites in the Toolchain: A Large-Scale Analysis of Attacks on the MCP Ecosystem

工具链中的寄生体:对MCP生态系统的大型分析

Shuli Zhao, Qinsheng Hou, Zihan Zhan, Yanhao Wang, Yuchong Xie, Yu Guo, Libo Chen, Shenghong Li, Zhi Xue

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract)

AI总结 研究发现MCP生态中存在寄生工具链攻击,攻击者通过恶意指令渗透工具链,窃取隐私数据,揭示MCP缺乏安全隔离机制,需加强防御。

Comments Accepted by IEEE Symposium on Security and Privacy, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.19622 2026-05-04 q-bio.NC 80%

Reassessing prediction in the brain: Pre-onset neural encoding during natural listening does not reflect pre-activation

重新评估大脑中的预测:自然听觉过程中预启动的神经编码不反映预激活

Sahel Azizpour, Britta U. Westner, Jakub Szewczyk, Umut Güçlü, Linda Geerligs

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract)

AI总结 研究通过MEG和ECoG验证了自然听觉中预启动信号的神经编码,发现其可能反映刺激依赖而非预激活,同时揭示了后预测的神经活动特征。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00197 2026-05-04 cs.MA cs.AI 79%

The $\textit{Silicon Society}$ Cookbook: Design Space of LLM-based Social Simulations

硅社会食谱:基于大语言模型的社会模拟设计空间

Aurélien Bück-Kaeffer, Sneheel Sarangi, Maximilian Puelma Touzel, Reihaneh Rabbany, Zachary Yang, Jean-François Godbout

机构 * McGill University(麦吉尔大学) Mila - Quebec Artificial Intelligence Institute(魁北克人工智能研究所) Université de Montréal(蒙特利尔大学) Ubisoft La Forge(育碧La Forge)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 研究如何通过大语言模型构建社会模拟网络,分析关键设计选择对模拟结果的影响,揭示设计空间的非平凡几何特性。

Comments 20 pages, 12 tables, under review at COLM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00480 2026-05-04 cs.CV 78%

Leveraging Vision-Language Models as Weak Annotators in Active Learning

利用视觉语言模型作为弱标注器在主动学习中的应用

Phuong Ngoc Nguyen, Kaito Shiku, Ryoma Bise, Seiichi Uchida, Shinnosuke Matsuo

机构 * Kyushu University(九州大学)

专题命中 其他LLM :language model(title,abstract)

AI总结 本文提出结合细粒度人工标注与粗粒度VLM生成弱标注的主动学习框架,通过实例级标注分配和系统噪声建模,在CUB200和FGVC-Aircraft数据集上优于现有方法。

Comments Accepted at ICIP2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00529 2026-05-04 cs.LG cs.AI cs.IR 73%

Hierarchical Abstract Tree for Cross-Document Retrieval-Augmented Generation

分层抽象树用于跨文档检索增强生成

Ziwen Zhao, Menglin Yang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出Ψ-RAG框架,通过迭代'合并与坍缩'过程构建分层抽象树索引,并引入多粒度检索代理,解决传统树状RAG在跨文档多跳问题中的分布适应性、结构隔离和抽象粗粒度等挑战,实验表明其在跨文档多跳问答任务中表现更优。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13511 2026-05-04 cs.CV cs.IR 67%

Adapting MLLMs for Nuanced Video Retrieval

为精细化视频检索适应大语言模型

Piyush Bagad, Andrew Zisserman

机构 * Visual Geometry Group, University of Oxford(牛津大学视觉几何组)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出一种统一嵌入模型,通过对比损失微调文本训练的多模态大语言模型,以处理视频检索中的时间、否定和多模态细微差别,实现最先进的检索性能。

Comments 38 Pages. Project page at http://bpiyush.github.io/tara-website

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00506 2026-05-04 cs.CL 57%

Surprisal Minimisation over Goal-directed Alternatives Predicts Production Choice in Dialogue

基于目标导向替代的惊奇最小化预测对话中的产出选择

Tom Utting, Mario Giulianelli, Arabella Sinclair

机构 * University of Aberdeen(阿伯丁大学) University College London(伦敦大学学院)

专题命中 其他LLM :language model(abstract);分类 cs.CL

AI总结 本文通过概率成本敏感选择模型,探讨对话中基于目标导向和上下文合理性的替代选择,发现基于目标导向替代的惊奇最小化在预测对话产出选择中表现最佳。

Comments 9 pages, to appear at ACL 2026 (Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00597 2026-05-04 cs.IR 50%

MUDY: Multi-Granular Dynamic Candidate Contextualization for Unsupervised Keyphrase Extraction

MUDY:面向无监督关键词提取的多粒度动态候选上下文化

Hyeongu Kang, Susik Yoon

专题命中 其他LLM :language model(abstract)

AI总结 MUDY提出一种以上下文为核心的框架,通过多粒度注意力机制和候选感知加权,提升关键词提取的局部上下文重要性识别能力,实验证明其在多个数据集上优于现有方法。

Comments Accepted to SIGIR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00081 2026-05-04 cs.CR cs.LO 50%

Alignment Contracts for Agentic Security Systems

代理安全系统的对齐契约

Isaac David, Marco Guarnieri, Arthur Gervais

专题命中 其他LLM :LLM(abstract)

AI总结 本文提出对齐契约框架,用于规范和强制执行可观测效果的行为约束,通过有限轨迹语义和模块化合同工程规则,确保安全系统的边界控制。

详情

展开后加载摘要…

URL PDF HTML 收藏