arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12083 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12083 篇

2606.21171 2026-06-23 cs.SE cs.AI 新提交 92%

An Exploratory Case Study of LLM-Assisted Refactoring and Gameplay Feature Generation in an Endless Runner Game

LLM辅助重构与无尽跑酷游戏玩法特征生成的探索性案例研究

Jan Wunderlich, Markus Kleffmann, Sebastian Lempert

机构 * IU International University of Applied Sciences(国际应用科学大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 通过GPT-4o在Python/Pygame无尽跑酷游戏中的案例,发现LLM在局部重构任务上表现可靠,但在需要跨系统交互的新玩法生成任务中成功率较低。

Comments 7 pages, 1 figure, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15788 2026-06-16 cs.CR cs.AI 新提交 92%

GAS-Leak-LLM: Genetic Algorithm-Based Suffix Optimization for Black-Box LLM Jailbreaking

GAS-Leak-LLM:基于遗传算法的后缀优化实现黑盒LLM越狱

Aman Anifer, Vignesh Kumar Kembu, Vishnu M, Antonino Nocera, Vinod P., Amal Murali PK, Akshay S Rajan

机构 * Department of Electrical, Computer and Biomedical Engineering(电气、计算机与生物医学工程系) University of Pavia(帕维亚大学) Department of Computer Applications(计算机应用系) Cochin University of Science and Technology(科钦科学技术大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出GAS-Leak-LLM方法,利用遗传算法在黑盒设置下自动进化对抗后缀以绕过LLM安全约束,验证了现有安全机制的不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.12247 2026-06-11 cs.CY cs.CL 新提交 92%

Beyond Third-Person Audits: Situated Interaction Auditing for User-Centered LLM Bias Research

超越第三人称审计:以用户为中心的LLM偏见研究的场景交互审计

Andrés Abeliuk, Cinthia Sanchez Macias, Valentina Alarcón, Álvaro Madariaga, Claudia Lopez

机构 * Department of Computer Science University of Chile Center for Artificial Intelligence (CENIA) Santiago, Chile(计算机科学系智利大学人工智能中心(CENIA)圣地亚哥,智利) Center for Artificial Intelligence (CENIA) Santiago, Chile(人工智能中心(CENIA)圣地亚哥,智利) Institute of Sociology Pontificia Universidad Católica de Chile Santiago, Chile(社会学研究所智利天主教大学圣地亚哥,智利)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 提出场景交互审计(SIA)框架,通过分析用户画像信号(如社会人口统计标记、写作风格和身份陈述)如何系统性地影响LLM响应质量、内容和语气,以用户为中心研究LLM偏见。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.08323 2026-06-09 cs.HC cs.AI 新提交 92%

"So There's a Catch-22 Here": How Early Adopters Who Build Multi-Agent LLM Systems Conceptualize Transparency

"所以这里有个第22条军规":构建多智能体LLM系统的早期采用者如何概念化透明度

Suchismita Naik, Samir Passi, Mihaela Vorvoreanu, Scott Saponas, Amanda Hall

机构 * Purdue University(普渡大学) Cornell University(康奈尔大学) Microsoft Research(微软研究院)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 通过访谈13位早期采用者,研究多智能体LLM系统构建者如何理解透明度,提出包含可重复性、调试、边界设定、可视化和审计的多维框架,强调透明度作为情境化的社会技术实践。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.30096 2026-05-29 cs.CR cs.AI 92%

How Reliable Are AI Attackers Against a Fixed Vulnerable Target? A 400-Run Empirical Study of LLM Penetration Testing Consistency

AI攻击者对固定脆弱目标的可靠性如何?LLM渗透测试一致性的400次运行实证研究

Galip Tolga Erdem

机构 * Independent Researcher(独立研究者)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 通过400次自主渗透测试运行(4个模型各100次),研究LLM在固定目标上攻击行为的一致性,发现模型间成功率差异显著且失败模式独特。

Comments 41 pages, 7 figures. Code and 400-run dataset: https://doi.org/10.5281/zenodo.20421592

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26506 2026-05-29 cs.CL cs.CR 92%

SafeReview: Defending LLM-based Review Systems Against Adversarial Hidden Prompts

SafeReview: 防御基于LLM的评审系统免受对抗性隐藏提示攻击

Yuan Xin, Yixuan Weng, Minjun Zhu, Ying Ling, Chengwei Qin, Michael Backes, Yue Zhang, Linyi Yang

机构 * CISPA Westlake University(西交利物浦大学) Southern University of Science and Technology(南方科技大学) HKUST (Guangzhou)(香港科技大学(广州))

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 提出SafeReview,一种共进化对抗训练框架,通过联合训练生成器和防御者模型,增强基于LLM的同行评审系统对对抗性隐藏提示的鲁棒性。

Comments 17 pages, 5 figures, 8 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.29062 2026-05-29 cs.CL 92%

Bosses, Kings, and the Commons: Cooperation Under Power Asymmetry in LLM Societies

老板、国王与公地:LLM 社会中权力不对称下的合作

Abhilekh Borah

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究通过引入不对称权力代理(老板或国王)的多智能体模拟框架 SovSim,发现权力不对称导致 LLM 社会中合作与可持续性严重崩溃,生存率较对称设置下降高达 87.3%。

Comments Paper under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.28632 2026-05-28 cs.CR cs.AI 92%

Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking

盲PRNG劫持:一种针对LLM水印的不可检测的完整性保持攻击

Ziyang You, Huilong He, Xiaoke Yang, Xuxing Lu

机构 * Fujian Provincial Key Laboratory of Automotive Electronics and Electric Drive(福建省汽车电子与电力驱动重点实验室) School of Electronic, Electrical and Physics(电子、电气与物理学院) Fujian University of Technology(福建理工大学) School of Humanities(人文学院) Institute of Applied Physics and Materials Engineering(应用物理与材料工程学院) University of Macau(澳门大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出SeedHijack攻击,通过替换伪随机数生成器(PRNG)在供应链层面对LLM水印进行盲攻击,同时保持完整性并规避检测。

Comments Preprint prepared for submission to IEEE TIFS. 12 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27913 2026-05-28 cs.LG 92%

Where LLM Annotators Fail: Label-Free Learning on Graphs with LLMs

LLM标注者失败之处:基于LLM的图上无标签学习

Safal Thapaliya, Jiatan Huang, Chuxu Zhang

机构 * University of Connecticut(康涅狄格大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 针对图节点分类中LLM标注噪声不仅依赖于类别还依赖于区域的问题,提出聚类感知噪声估计(CANE)框架,通过估计聚类条件可靠性来筛选和校正伪标签,在多个图基准上超越现有无标签方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24421 2026-05-26 cs.CR cs.LG 92%

Poisoning the Watchtower: Prompt Injection Attacks Against LLM-Augmented Security Operations Through Adversarial Log Content

毒害瞭望塔:通过对抗性日志内容对LLM增强的安全运营进行提示注入攻击

Rohan Pandey, Archit Bhujang

机构 * DigitalOcean Arizona State University(亚利桑那州立大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 研究攻击者控制的日志字段如何向LLM注入指令(日志基底提示注入),提出四类攻击分类,并评估不同防御下的攻击成功率。

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.23867 2026-05-25 cs.HC cs.AI 92%

Human Decision-Making with Persuasive and Narrative LLM Explanations

具有说服性和叙事性LLM解释的人类决策

Laura R. Marusich, Mary Grace Kozuch Dhooghe, Jonathan Z. Bakdash, Murat Kantarcioglu

机构 * DEVCOM Army Research Laboratory(美国陆军研发实验室) University of Texas at Dallas(德克萨斯大学达拉斯分校) Virginia Polytechnic Institute and State University(弗吉尼亚理工大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 通过大规模行为实验,研究LLM生成的叙事解释的说服力程度对人类决策准确性的影响,发现叙事解释未显著提升决策准确性,但增加了对AI的依赖,且高说服力叙事可能损害决策反应时间和辨别能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15949 2026-05-21 q-fin.TR cs.AI 92%

ATLAS: Adaptive Trading with LLM AgentS Through Dynamic Prompt Optimization and Multi-Agent Coordination

ATLAS:通过动态提示优化和多智能体协调实现LLM智能体的自适应交易

Charidimos Papadakis, Angeliki Dimitriou, Giorgos Filandrianos, Maria Lymperaiou, Konstantinos Thomas, Giorgos Stamou

机构 * School of Electrical and Computer Engineering, AILS Laboratory(电气与计算机工程学院,AILS实验室) National Technical University of Athens(雅典国家技术大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出ATLAS框架,通过动态提示优化和多智能体协调,解决LLM在金融交易中的适应性问题,提升交易决策的鲁棒性和执行效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.17787 2026-05-19 cs.LG 92%

Revisiting the Adam-SGD Gap in LLM Pre-Training: The Role of Large Effective Learning Rates

重新审视LLM预训练中Adam与SGD的差距:大有效学习率的作用

Athanasios Glentis, Dawei Li, Chung-Yiu Yau, Mingyi Hong

机构 * University of Minnesota(明尼苏达大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文通过实证和理论分析,发现SGD在LLM预训练中表现较差的原因在于其无法维持与Adam相媲美的有效学习率,而大有效学习率需求源于小梯度范数和大权重-梯度比,且在大批次大小下更加明显。通过简单剪枝机制,SGD在大学习率下能恢复大部分Adam性能,实验显示验证损失差距从超过50%降至约3.5%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.04932 2026-05-19 cs.CL 92%

Beyond the Final Actor: Modeling the Dual Roles of Creator and Editor for Fine-Grained LLM-Generated Text Detection

超越最终作者:为细粒度LLM生成文本检测建模创作者与编辑的双重角色

Yang Li, Qiang Sheng, Zhengjia Wang, Yehan Yang, Danding Wang, Juan Cao

机构 * ict.ac.cn(中国科学院)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出RACE方法,通过建模创作者和编辑的双重角色,实现细粒度LLM生成文本检测,以更精确地区分不同类型的文本,从而为LLM监管提供政策对齐的解决方案。

Comments ACL 2026 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21410 2026-05-12 stat.ML cs.LG 92%

Learning When to Trust LLM Priors: A Validated Framework for Semantic Prior Integration

学习何时信任LLM先验:一种经过验证的语义先验整合框架

Erica Zhang, Naomi Sagan, Danny Tse, Fangzhao Zhang, Mert Pilanci, Jose Blanchet

机构 * Stanford University School of Engineering(斯坦福大学工程学院)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 Statsformer框架通过验证LLM生成的语义先验,提升监督统计学习的可靠性,有效整合不同模型的先验信息,自动降低不可靠先验的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27006 2026-05-01 cs.SE cs.AI 92%

Beyond Accuracy: LLM Variability in Evidence Screening for Software Engineering SLRs

超越准确性:LLM在软件工程SLR证据筛选中的变异性

Gilberto Sussumu Hida, Danilo Monteiro Ribeiro, Erika Yahata

机构 * AIBL CESAR School(AIBL CESAR学校) CMCC UFABC

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究评估了LLM在证据筛选中的性能和变异性,分析了输入元数据对性能的影响,并比较了LLM与传统分类器的实证效果。

Comments 16 pages, 12 figures. Earlier, shorter, conference-style version of a more comprehensive journal manuscript currently under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.07387 2026-04-30 cs.AR cs.AI 92%

A Self-Calibrating Framework for Analog Circuit Sizing Using LLM-Derived Analytical Equations

基于LLM衍生分析方程的自校准模拟电路尺寸设计框架

Antonio J. Bujana, Aydin I. Karsilayan

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种自校准的模拟电路尺寸设计框架,通过LLM生成可追溯设计原理的分析方程,结合确定性校准循环和预测误差反馈机制,实现跨工艺节点的自动校准。

Comments 14 pages, 4 figures, 9 tables. V2: Extended to 5 topology families (8-30 transistors), 3 process nodes, and quantitative comparison against 4 published methods

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13400 2026-04-29 cs.CY cs.AI 92%

Justice in Judgment: Unveiling (Hidden) Bias in LLM-assisted Peer Reviews

审判中的公正:揭示(隐藏)在LLM辅助同行评审中的偏见

Sai Suresh Macharla Vasu, Ivaxi Sheth, Hui-Po Wang, Ruta Binkyte, Mario Fritz

机构 * CISPA Helmholtz Center for Information Security(CISPA 欧洲信息安全研究中心) Saarland University(萨尔兰大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究LLM生成的同行评审中的偏见,通过操控作者元数据发现机构偏见、资历偏好和性别影响,揭示隐性偏见在软评分中的显现。

Comments Findings of ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.24444 2026-04-28 cs.CL 92%

Can You Make It Sound Like You? Post-Editing LLM-Generated Text for Personal Style

你能让它听起来像你吗?为个人风格进行后编辑LLM生成文本

Connor Baumler, Calvin Bao, Huy Nghiem, Xinchen Yang, Marine Carpuat, Hal Daumé

机构 * III University of Maryland(III 马里兰大学)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究探讨用户通过后编辑LLM生成文本来塑造个人风格的效果,发现后编辑提高了文本与用户自身写作的相似性,但降低了与LLM生成文本的相似性,同时减少了风格多样性。

Comments ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.02861 2026-04-27 cs.DB cs.AI 92%

LLM+Graph@VLDB'2025 Workshop Summary

LLM+Graph@VLDB'2025研讨会总结

Yixiang Fang, Arijit Khan, Tianxing Wu, Da Yan, Shu Wang

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文总结了LLM与图结构数据结合的研究进展,探讨了算法、系统及图机器学习在实际应用中的挑战与创新解决方案。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.19440 2026-04-22 cs.CL cs.NE 92%

What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search

什么使一个LLM成为好的优化器?LLM引导的进化搜索轨迹分析

Xinhao Zhang, Xi Chen, François Portet, Maxime Peyrard

机构 * Univ. Grenoble Alpes, CNRS, Grenoble INP, LIG(格勒诺布尔阿尔卑斯大学、国家科学研究中心、格勒诺布尔INP、LIG实验室)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究通过分析15个LLM在8个任务上的优化轨迹,发现初始能力相似的模型会产生差异大的搜索轨迹,强优化器表现为局部细化,弱优化器则出现语义漂移,解决方案新颖性只有在搜索保持局部化时才有效。

Comments 9 pages, 8 figures, Accepted at Findings of ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14808 2026-04-17 cs.CL 92%

Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem

将LLM去学习视为一个非对称的双任务学习问题

Zeguan Xiao, Siqing Li, Yong Wang, Xuetao Wei, Jian Yang, Yun Chen, Guanhua Chen

机构 * Shanghai University of Finance and Economics(上海金融学院) Alibaba Group(阿里巴巴集团) Southern University of Science and Technology(南方科技大学) Beihang University(北航) MoE Key Laboratory of Interdisciplinary Research of Computation and Economics(计算与经济交叉学科研究教育部重点实验室)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文将LLM去学习视为非对称双任务问题,提出优先保留的梯度合成框架,通过分离任务特定梯度提取与冲突感知组合,改进梯度冲突解决方法,实验证明通过重塑梯度几何而非重新平衡损失,有效缓解去学习-保留的权衡。

Comments ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11312 2026-04-16 cs.SI cs.AI cs.CY cs.MA physics.soc-ph 92%

Network Effects and Agreement Drift in LLM Debates

神经网络效应与LLM辩论中的共识漂移

Erica Cau, Andrea Failla, Giulio Rossetti

机构 * Department of Computer Science, University of Pisa(比萨大学计算机科学系) ISTI-CNR(意大利国家研究委员会ISTI研究所)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文通过受控同质性和群体规模的网络生成模型,研究了LLM代理在多轮辩论中的集体行为,揭示了共识漂移现象,强调需区分结构效应与模型偏差以正确理解LLM群体的行为。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12922 2025-12-16 cs.LG 92%

LLM-based Personalized Portfolio Recommender: Integrating Large Language Models and Reinforcement Learning for Intelligent Investment Strategy Optimization

基于大语言模型的个性化投资组合推荐器:整合大语言模型与强化学习以实现智能投资策略优化

Bangyu Li, Boping Gu, Ziyang Ding

专题命中 其他LLM :LLM(title,abstract);large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 本文提出一种结合大语言模型与强化学习的个性化投资组合推荐系统,旨在通过智能优化提升投资策略的适应性和有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06874 2025-09-08 cs.HC cs.AI 92%

LLM-D12: A Dual-Dimensional Scale of Instrumental and Relational Dependencies on Large Language Models

Ala Yankouskaya, Areej B. Babiker, Syeda W. F. Rizvi, Sameha Alshakhsi, Magnus Liebherr, Raian Ali

专题命中 其他LLM :LLM(title,abstract);large language model(title,abstract);language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14869 2025-08-21 q-bio.NC cs.CL 92%

The Prompting Brain: Neurocognitive Markers of Expertise in Guiding Large Language Models

Hend Al-Khalifa, Raneem Almansour, Layan Abdulrahman Alhuasini, Alanood Alsaleh, Mohamad-Hani Temsah, Mohamad-Hani_Temsah, Ashwag Rafea S Alruwaili

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(title);LLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23566 2025-04-01 cs.CL 92%

When LLM Therapists Become Salespeople: Evaluating Large Language Models for Ethical Motivational Interviewing

Haein Kong, Seonghyeon Moon

专题命中 其他LLM :LLM(title,abstract);large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06471 2025-01-14 cs.AI 92%

The Internet of Large Language Models: An Orchestration Framework for LLM Training and Knowledge Exchange Toward Artificial General Intelligence

Wilson Wei, Nicholas Chen, Yuxuan Li

专题命中 其他LLM :LLM(title,abstract);large language model(title,abstract);language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21414 2024-09-25 eess.AS cs.CL 92%

Towards interfacing large language models with ASR systems using confidence measures and prompting

Maryam Naderi, Enno Hermann, Alexandre Nanchen, Sevada Hovsepyan, Mathew Magimai. -Doss

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(title,abstract);分类 cs.CL

Comments 5 pages, 3 figures, 5 tables. Accepted to Interspeech 2024

Journal ref Proc. Interspeech 2024, 2980-2984

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08147 2024-09-13 cs.CL 92%

LLM-POTUS Score: A Framework of Analyzing Presidential Debates with Large Language Models

Zhengliang Liu, Yiwei Li, Oleksandra Zolotarevych, Rongwei Yang, Tianming Liu

专题命中 其他LLM :LLM(title,abstract);large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏