arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 693 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 693 篇

2608.01538 2026-08-04 cs.NI cs.SY eess.SY 新提交 91%

From Network Automation to Trustworthy Autonomous Networking in the LLM Era: A Network Control Intelligence Perspective

从网络自动化到大语言模型(LLM)时代的可信自主网络:网络控制智能视角

Tianzhu Zhang, Changgang Zheng, Shanshan Wang, Yarui Zhang, Lina Shi, Yue Jin, Xiaofei Wang, Meikang Qiu

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 本文以网络控制智能(NCI)为框架,梳理网络控制系统三阶段演进,提出可信自主网络定义及参考架构,明确LLM赋能运营的集成模式与相关研究议程。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30908 2026-07-01 cs.CY 新提交 91%

Demystify, Use, Reflect, Assess (DURA): An Experience Report on LLM Integration in CS2

揭秘、使用、反思、评估(DURA):CS2课程中集成LLM的经验报告

Margaret Ellis, Nikitha Donekal Chandrashekar, Sehrish Basir Nizamani, Mohammed Farghally, Jake O'Brien, Naren Ramakrishnan

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 提出DURA框架(揭秘-使用-反思-评估),重构CS2课程以允许使用LLM,通过揭秘、指导使用、反思和增加监考评估,促进学生元认知,学生报告了LLM的多种用途并提高了对教师关怀的感知。

Comments 7 pages, 2 figures, 6 tables. Experience report accepted to SIGCSE Virtual 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26505 2026-06-26 cs.SE cs.HC 新提交 91%

Same Scrutiny, More Time: Eye Tracking Insights into Reviewing LLM-Labelled Code

同样审查,更多时间:眼动追踪洞察审查LLM标记代码

Ranim Khojah, Francisco Gomes de Oliveira Neto, Mazen Mohamad, Julian Frattini, Philipp Leitner

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 通过Wizard-of-Oz实验和眼动追踪,发现开发者审查LLM标记代码时更专注但策略调整,提示需重新审视AI政策。

Comments Accepted at the 41st IEEE/ACM International Conference on Automated Software Engineering (ASE 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.21349 2026-06-24 cs.CR 新提交 91%

LLM-assisted Generation of Pseudo-C2 Servers for IoT Malware Dynamic Analysis

LLM辅助生成伪C2服务器用于物联网恶意软件动态分析

K. Hasui, S. Matsugaya, M. Shimamura, M. Hashimoto

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 提出结合Ghidra与大语言模型(LLM)的系统,从恶意软件二进制中提取通信规范并自动生成伪C2服务器,在Mirai上实现100%规范提取并复现7/10种DDoS攻击。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.23343 2026-06-23 cs.CY 新提交 91%

Group Selection Promotes Prosocial Prompts in Populations of LLM Agents

群体选择促进LLM智能体群体中的亲社会提示

Luis Celiktemel, Edward Eichhorn, Levin Brinkmann, Robin Schimmelpfennig, Aron Vallinder, Yaomin Jiang, Edward Hughes, Iyad Rahwan

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 研究通过多智能体模拟框架,发现群体选择机制能促进LLM智能体群体中的亲社会提示演化,稳定合作行为,而个体选择则导致集体背叛。

Comments 23 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.19644 2026-06-19 cs.SE 新提交 91%

Prompt Quality and Pull Request Outcomes: A Stage-Based Empirical Study of LLM-Assisted Development

提示质量与拉取请求结果:基于阶段的LLM辅助开发实证研究

Richard Sserunjogi, Daniel Ogenrwot, John Businge

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 通过分析265个开发者与ChatGPT的交互,研究提示结构(上下文、具体性、验证)对LLM辅助开发中代码生成、采纳和集成深度的影响,发现不同维度在不同阶段有不同作用。

Comments 48 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15267 2026-06-16 eess.AS cs.SD 新提交 91%

Dynamic Prosody Prediction in LLM-based TTS for Improving Speaker Similarity

基于LLM的TTS中的动态韵律预测以提高说话人相似度

Zhenwei Mou, Liping Chen, Yajun Hu, Zhen-Hua Ling, Xin Fang, Jianqing Gao

机构 * University of Science and Technology of China(中国科学技术大学) iFLYTEK

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 针对LLM-based TTS忽略风格特定韵律模式导致说话人相似度不足的问题,提出基于先前预测语音的动态音节韵律预测方法,显著提升韵律学习能力和说话人相似度。

Comments Accepted to INTERSPEECH 2026. 5 pages, 2 figures. Audio samples: https://muzw.github.io/dynapros/

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15264 2026-06-16 eess.AS cs.SD 新提交 91%

DuraMark: Duration-Embedded Watermarking in LLM-based TTS

DuraMark: 基于LLM的文本转语音中的时长嵌入水印

Zhenwei Mou, Weili Jiang, Liping Chen, Zhen-Hua Ling, Kong Aik Lee, Kai Gao, Boyu Zhao

机构 * University of Science and Technology of China(中国科学技术大学) Institute of Forensic Science, Ministry of Public Security(公安部刑侦科学研究所) The Hong Kong Polytechnic University(香港理工大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 提出DuraMark,一种基于音节时长编辑的信息级水印框架,利用可控时长的LLM-TTS模型嵌入水印,并使用时长提取器检测,有效抵抗生成式攻击。

Comments Accepted to INTERSPEECH 2026. 5 pages, 1 figure. Audio samples: https://muzw.github.io/duramark_demo/

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10853 2026-06-10 eess.AS 新提交 91%

Speech Encoder Fusion for LLM-based Automatic Speech Recognition

面向基于LLM的自动语音识别的语音编码器融合

Jakob Poncelet, Hugo Van hamme

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 研究融合多个预训练语音编码器以增强基于LLM的ASR性能,提出多种融合策略并在多场景下验证其有效性。

Comments Accepted at Interspeech 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24504 2026-06-24 cs.AI 新提交 91%

On the Smallness of the Large Language Models Scaling Exponents

论大型语言模型缩放指数的小值问题

Sauro Succi, Peter V. Coveney, Alex Hansen

机构 * Italian Institute of Technology(意大利理工学院) PoreLab, Physics Department, Norwegian University of Science and Technology(挪威科技大学物理系PoreLab实验室) Centre for Computational Science, Chemistry Department, University College of London(伦敦大学学院化学系计算科学中心)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(summary_cn,abstract_cn);分类 cs.AI

AI总结 本文讨论当前LLM缩放指数小导致能源不可持续的问题,指出忽略无限数据下损失函数非零的“基座效应”无法解决该问题,并基于流体湍流现象学模型类比数据平滑性对缩放指数的影响。

Comments 11 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22009 2026-06-23 cs.CL eess.AS 新提交 91%

Benchmarking Large Language Models for Grapheme-to-Phoneme Conversion: A Japanese Case Study

大型语言模型在字素到音素转换中的基准测试:以日语为例

Tomoki Koriyama

机构 * CyberAgent, Japan(日本CyberAgent公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);prompting(abstract)

AI总结 本研究以日语为例,对30多种大型语言模型进行字素到音素转换基准测试,发现模型大小、版本和日语专项训练是关键因素,最佳模型字符错误率低于0.52%,优于传统工具。

Comments accepted to Interspeech 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.28667 2026-08-03 cs.LG cs.AI math.DS 新提交 91%

Guarantees on Dynamical System Distinguishability for LLM Token Generation

大语言模型(LLM)令牌生成的动力系统可区分性保证

Mohamed Akrout, Dan Wilson

机构 * University of Tennessee(田纳西大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 该研究将LLM响应分类任务形式化为随机线性DS的二元假设检验,证明基于DS的分类误判率随序列长度指数衰减,还建立跨嵌入泛化的可迁移可区分性下界,解释了该方法的经验性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.27134 2026-07-30 cs.AI cs.CL cs.GT 新提交 91%

Linguistic Monoculture in LLM-Assisted Language Use

LLM辅助语言使用中的语言单一文化

Suhas Thejaswi, Juhi Kulshreshta, Lutz Oettershagen

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 该研究针对LLM辅助语言使用中的语言单一文化问题,构建数学框架分析作者与LLM的共同演化机制,发现个性化可保留语言多样性,且个体理性作者的过度一致性会产生负外部性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30668 2026-07-01 cs.NE cs.AI cs.CL cs.MA nlin.AO q-bio.PE 新提交 91%

Emergent Culture in Minimal LLM Systems

最小LLM系统中的涌现文化

Simon Jones, Sabine Hauert

机构 * University of Bristol(布里斯托大学)

专题命中 其他LLM :LLM(title,title_cn);prompting(abstract);分类 cs.CL、cs.AI

AI总结 研究在极简条件下LLM智能体如何自发合作并产生复杂文化产物,通过动态系统分析揭示超越熵视界的结构化长程相干性。

Comments 9 pages, 6 figures. Accepted for publication at Alife 2026 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.28090 2026-06-29 cs.HC 新提交 91%

Typing Behavior in Human-LLM Interaction: Keystroke Dynamics Reveal Cognitive Effort During Prompting

人机交互中的打字行为:击键动态揭示提示过程中的认知努力

Laura Schütz, Yousri Cherif, Clara Sayffaerth, Thomas Weber, Francesco Chiossi

专题命中 其他LLM :LLM(title,abstract);prompting(title,abstract);large language model(abstract);language model(abstract)

AI总结 本研究通过击键动态分析用户在与大语言模型交互时的认知努力,发现困难任务导致更多击键、更慢打字和更多停顿,但击键无法预测感知的输出有用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.28651 2026-08-03 cs.HC cs.CL cs.CY 新提交 91%

Measuring Cognitive Engagement in Collaborative Discourse with an Extended ICAP Framework: Comparing Human Annotation, In-Context Learning, and Reflective LLM Agents

基于扩展ICAP框架的协作对话认知投入度测量:人类标注、上下文学习与反思型大语言模型智能体的比较

Lan Anh Do, Hanling Jiang, Shuchin Aeron, Ayanna K. Thomas

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本研究用扩展7点ICAP框架测量协作对话认知投入度,对比人类标注、ICL及反思型LLM智能体,发现人类标注信度更高,智能体方法具潜力,需强化人类标注与LLM方法的交互。

Comments Accepted as a full paper in the CogSci 2026 proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.27373 2026-07-31 cs.CR cs.AI 新提交 91%

RoguePrompt: Dual-Layer Encoding for Self-Reconstruction to Circumvent LLM Moderation

RoguePrompt:用于自重构以规避大型语言模型(LLM)审核的双层编码

Benyamin Tafreshian, Prathamesh Dhake

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 RoguePrompt是一种采用双层编码(维吉尼亚密码+ROT13)的LLM越狱流程,在黑盒威胁模型下对313个被拒提示词测试,实现93.93%的过滤绕过率,提供多阶段越狱失效的阶段级证据。

Comments This manuscript supersedes the preliminary version available as arXiv:2511.18790. The work has been substantially revised, expanded, and reorganized, with a refined threat model, revised methodology, clearer stage-level evaluation criteria, and expanded analysis of moderation bypass, instruction reconstruction, and execution

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.23026 2026-06-23 cs.AI 新提交 91%

A Stackelberg Framework for Resource-Aware LLM Agents: Learning, Repair, and Conditional Guarantees

面向资源感知的LLM智能体的Stackelberg框架:学习、修复与条件保证

Baoxun Wang

机构 * Platform and Content Group, Tencent(腾讯平台与内容事业群)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 提出基于Stackelberg博弈的资源治理框架,通过控制器设定质量目标与成本激励,执行器响应资源动作,结合学习、修复与条件保证,在300轮实验中平均token成本降低17.4%,质量无显著差异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.10703 2026-08-12 cs.LG cs.AI cs.CL cs.HC 新提交 91%

Your LLM, Your Style: Behavioral Mode Axes for LLM Behavioral Control

你的大语言模型,你的风格:用于大语言模型行为控制的行为模式轴

Haoze Liu, Run Liu, Haiying Xu, Jiahui Han, Siyuan Fang, Siyu Yan, Huiqi Deng, Guanchu Wang, Na Zou

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本研究提出情境化行为数据框架,构建3200个对比性行为场景,发现LLMs有稳定且模型特异性的行为特征,提出行为模式轴控制LLM行为,表明其类人格倾向是可测量可控的行为模式。

Comments 33 pages, 8 figures. Code and data: https://github.com/lhz191/LLM-Behavioral-Personality

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20041 2026-06-19 econ.GN cs.AI cs.LG q-fin.EC q-fin.GN 新提交 91%

AI Economist Agent: An Agentic Framework for Model-Grounded Economic Analysis with RAG, Knowledge Graphs, and Large Language Models

AI经济学家代理:一种基于模型的经济分析代理框架,结合RAG、知识图谱和大语言模型

Masahiro Kato

机构 * Mizuho-DL Financial Technology, Co., Ltd.(Mizuho-DL金融科技有限公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.AI、cs.LG

AI总结 提出一种基于RAG的AI经济学家代理框架,利用知识图谱和大语言模型进行经济情景分析,通过代理规划、检索证据、选择模型并生成报告,提高经济叙事的连贯性和可追溯性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.17612 2026-06-17 cs.SE 新提交 90%

PracRepair: LLM-Empowered Automated Program Repair Inspired by Human-Like Debugging Practices

PracRepair: 受人类调试实践启发的大语言模型赋能自动化程序修复

Yu Cheng, Zhongxin Liu, Zhenchang Xing, Chao Ni, Qing Huang, Xiaoxue Ren

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);foundation model(abstract)

AI总结 提出PracRepair框架,通过构建按需静态-动态上下文、进行问题驱动的故障诊断并迭代细化补丁,利用动态信息提升LLM在程序修复中的效果,在Defects4J和真实世界漏洞上取得最优性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.12719 2026-08-14 cs.GT cs.AI 新提交 90%

Error-Aware Reverse Auction Mechanism for Large Language Model Routing

面向大语言模型路由的误差感知反向拍卖机制

Haolong Chen, Zhengyuan Xin, Liang Zhang, Lei Xue, Guangxu Zhu

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.AI

AI总结 针对大语言模型路由的信息风险不匹配与可扩展性瓶颈,提出EA-RAM反向拍卖机制,证明其相关特性,实验显示该机制鲁棒且性能优于集中式基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.09551 2026-08-11 cs.CL 新提交 90%

Pragmatic Attack Surface: Vulnerabilities of Implicit Context in Large Language Models

语用攻击面:大语言模型中隐式上下文的漏洞

Bocheng Chen, Han Zi, Roucheng Ou, Yawei Liu, Minyue Chen, Zimo Qi, Rongrong Wang, Guangliang Liu

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.CL

AI总结 本文提出大语言模型存在语用攻击面漏洞,利用人类语言解读与安全对齐的隐式上下文不匹配,所提攻击方法在各类模型上的成功率大幅优于基线方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.04999 2026-08-06 eess.SY cs.AI cs.SY 新提交 90%

ORACLE: A Multi-Objective Reinforcement Learning-Based Analog Circuit Design Optimizer with Large Language Models-Guided Exploration

ORACLE:一种基于多目标强化学习且结合大语言模型引导探索的模拟电路设计优化器

Osei Brempong, Mohammed Ayman Habib, Vivan Poddar, Morteza Fayazi

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.AI

AI总结 ORACLE是结合多目标强化学习与大语言模型引导探索的开源模拟电路设计优化框架,可无需重训即生成多权衡设计,运行时间大幅缩短且指标满足度与品质因数优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.00538 2026-08-04 cs.CL 新提交 90%

DE-NER : Zero-shot Named Entity Recognition via Dialogue Elicitation of Large Language Models

DE-NER:通过大型语言模型的对话引导实现零样本命名实体识别

Xuankang Zhang, Jiangming Liu

机构 * Yunnan University(云南大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.CL

AI总结 DE-NER是一个对话引导框架,利用大型语言模型的对话能力解决零样本命名实体识别的提示与演示工程局限,在多基准零样本设置下平均F1值提升3.75%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.05761 2026-07-08 cs.AI 新提交 90%

Synthetic Consumer Insight Generation with Large Language Models

使用大语言模型生成合成消费者洞察

Stephen L. France, Pia. A. Albinsson

机构 * Mississippi State University(密苏里州立大学) Appalachian State University(阿巴拉契亚州立大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

AI总结 研究探讨大语言模型能否为投射技术生成合成消费者数据,通过在多任务等方面测试其生成的回答,并与人类回答比较分析。结果显示二者在主题联想上有重叠但也有差异,还给出利用大语言模型生成数据的建议及对其局限性的认识。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15565 2026-06-16 cs.HC cs.LG 新提交 90%

If These Walls Could Talk: Critical Play with Large Language Models in Museums

如果这些墙会说话:博物馆中大语言模型的批判性游戏

Anders Sundnes Løvlie

机构 * The Dalí Museum(达利博物馆)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.LG

AI总结 针对博物馆中大语言模型聊天机器人不可靠但吸引人的矛盾,提出设计批判性游戏,将机器人作为虚构角色呈现历史叙事、话语风格和多元视角。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.12713 2026-08-14 cs.CR cs.AI cs.CL 新提交 90%

Tracing Provenance and Detecting Tampering with Complementary LLM Watermarks

追踪来源并检测互补大语言模型水印的篡改

Xiaoyan Feng, Yanjun Zhang, He Zhang, Leo Yu Zhang, Shirui Pan

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究提出一种互补LLM水印方法,通过嵌入鲁棒与脆弱双信号,实现来源追踪与篡改检测,在两类LLM和两类提示数据集上,其篡改检测率优于现有方法,且保持了良好的归属鲁棒性与困惑度。

Comments 11 pages, 7 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.11513 2026-08-13 cs.SE cs.AI cs.CL 新提交 90%

Do Influence Tactics Matter? Investigating Prompt Framing Effects in LLM Code Generation

影响策略重要吗?研究大语言模型代码生成中的提示词框架效应

Alex Deaconu, Anubhav Gupta, Manaal Basha, Nicholas Haydu, Gema Rodríguez-Pérez

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究首次大规模实证探究基于心理学的影响策略诱导的提示词框架对LLM代码生成的影响,发现强调紧迫性的框架会降低代码正确性与安全性,为设计人机交互提供实践见解。

Comments Accepted for publication in Empirical Software Engineering. This is the accepted manuscript version. 37 pages, 3 figures

Journal ref Empirical Software Engineering 32 (2026) 13

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.06992 2026-08-10 cs.CL cs.AI cs.DB 新提交 90%

GPTKB 2.0: Browsing, Querying, and Auditing a Disambiguated LLM-Derived Knowledge Base

GPTKB 2.0:浏览、查询与审计经消歧的大语言模型衍生知识库

Yujia Hu, Tuan-Phong Nguyen, Simon Razniewski

机构 * ScaDS.AI Dresden/Leipzig(ScaDS.AI德累斯顿/莱比锡) TU Dresden(德累斯顿工业大学) Institute for AI, VNU University of Engineering and Technology(越南国家大学工程技术学院人工智能研究所)

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 该研究推出了可浏览、查询、审计的GPTKB 2.0系统,其为经上下文引导消歧的LLM衍生知识库,规模达3840万条三元组,支持多种查询及实体链接功能,且提供网络演示与离线下载。

Comments 7 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏