arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12096 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12096 篇

2408.10268 2025-11-19 cs.SE cs.AI cs.LG 90%

Generating Streamlining Constraints with Large Language Models

Florentina Voboril, Vaidyanathan Peruvemba Ramaswamy, Stefan Szeider

机构 * Algorithms and Complexity Group TU Wien(算法与复杂性组维也纳技术大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

Comments 23 page; deeper analysis of streamliners and statistics about benchmark instances added

Journal ref F. Voboril, V. P. Ramaswamy, S. Szeider, Generating Streamlining Constraints with Large Language Models, Journal of Artificial Intelligence Research, volume 84, pages 16:1-16:19, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.16891 2024-04-29 cs.CR cs.AI cs.CL cs.CY 90%

Attacks on Third-Party APIs of Large Language Models

Wanru Zhao, Vidit Khazanchi, Haodi Xing, Xuanli He, Qiongkai Xu, Nicholas Donald Lane

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

Comments ICLR 2024 Workshop on Secure and Trustworthy Large Language Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27152 2026-08-14 cs.SI 版本更新 90%

Disrupting Networks: Amplifying Social Dissensus via Opinion Perturbation and Large Language Models

干扰网络:通过观点扰动和大语言模型放大社会分歧

Erica Coppolillo, Giuseppe Manco

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 该研究基于Friedkin-Johnsen模型,利用强化学习框架微调大语言模型生成干扰性文本,实现了对社交网络的战略性干扰,其结果对内容审核等领域有重要参考价值。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.10613 2026-08-12 cs.SE 新提交 90%

CausalRepair: Bridging the Causality Gap in Large Language Model-Based Automated Program Repair via Dual-Slicing

CausalRepair:通过双重切片弥合基于大语言模型的自动化程序修复中的因果差距

Linhao Wu, Yizhou Chen, Zhen Yang, Pengyu Xue, Dan Hao

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 CausalRepair是基于最小因果上下文的对话驱动型自动化程序修复框架,通过双重切片策略构建因果相关上下文,在Defects4J数据集上修复313个漏洞,性能优于现有方法且修复成本更低。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.07488 2026-08-11 cs.HC cs.CY 新提交 90%

Large Language Models Explain Experts Better Than Experts Themselves

大型语言模型(LLM)比专家自身更擅长解释专家

Mina Cho, Russell J. Funk, Alok Gupta, Mochen Yang

专题命中 其他LLM :LLM(title_cn,abstract);large language model(title);language model(title)

AI总结 该研究表明,大型语言模型可从专家行为中外部化隐性知识,提升决策质量并帮助新手接近专家表现,为波兰尼悖论提供实证支持,凸显其作为克服专家表述瓶颈的可扩展工具的潜力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22356 2026-08-07 cs.HC 版本更新 90%

Large Language Model Counterarguments in Older Adults: Cognitive Offloading or Susceptibility to Moral Persuasion?

大语言模型在老年人中的反论点:认知卸载还是对道德说服的脆弱性?

Kou Tamura, Sayaka Ishibashi, Ayana Goma, Kenta Yamamoto, Kouhei Masumoto

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 研究发现大语言模型能显著影响老年人和年轻人的道德判断,老年人更易受说服,认知功能较低者在情感冲突困境中更易接受反论点,表明LLMs可能成为认知卸载工具,但也对认知脆弱者构成风险。

Comments This paper has been published in Computers in Human Behavior. The final published version is available at https://doi.org/10.1016/j.chb.2026.109142

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02420 2026-08-04 cs.HC cs.CY 新提交 90%

WIP: Chat-Debugging: Large Language Model as a Hardware Debugging Assistant

WIP:Chat-Debugging:大型语言模型作为硬件调试助手

Andrew Ash, John Hu

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 该研究提出将大型语言模型应用于硬件调试的Chat-Debugging方法,通过人机交互提升电气类学生的调试信心与技能,填补了物理硬件调试辅助工具的研究空白。

Comments This is the accepted version of a paper accepted for presentation at the 2026 IEEE Frontiers in Education Conference (FIE). The final version will be available via IEEE Xplore at: https://ieeexplore.ieee.org/Xplore/home.jsp

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.22657 2026-07-30 cs.NE 版本更新 90%

Automatic programming via large language models with population self-evolution for dynamic fuzzy job shop scheduling problem

结合种群自进化大语言模型的动态模糊作业车间调度问题自动编程

Jin Huang, Qihao Liu, Xinyu Li, Liang Gao, Yue Teng

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 该研究针对动态模糊作业车间调度问题,提出种群自进化框架结合大语言模型自动设计启发式调度规则,性能优于多种现有方法。

Comments 13 pages, 10 figures. Accepted for publication in IEEE Transactions on Fuzzy Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.03202 2026-07-21 cs.CY 版本更新 90%

Prosocial Persuasion at Scale? Large Language Models Outperform Humans in Donation Appeals Across Levels of Personalization

大规模的亲社会说服?大语言模型在不同个性化水平上的捐赠呼吁中优于人类

John Caffier, Olga Stavrova, Bennett Kleinberg

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 研究探讨大语言模型生成的捐赠呼吁在不同个性化水平下的有效性,发现其在捐款数额、参与度和说服力上均优于人类创作内容,但虚假个性化会带来负面影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.12235 2026-07-15 cs.CY 新提交 90%

A Semi-Automated System for Generating Dialogue-Based TTS Lessons Using Large Language Models: An Exploratory Study of Educational Potential

一种使用大语言模型生成基于对话的语音合成课程的半自动系统:教育潜力的探索性研究

Gendo Kumoi, Fumie Watanabe, Tota Suko, Takashi Ishida, Yuko Kuma, Manabu Kobayashi, Shigeichi Hirasawa

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 研究提出用大语言模型和文本转语音技术生成基于对话课程的半自动系统,经准实验探索其教育潜力。系统通过三阶段工作流程增强教育工作者,引入新方法。实验表明对话TTS在理解等方面优于单声道TTS,为TTS音频教育可接受性及课程形式设计提供依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.28127 2026-06-29 cs.CL cs.AI cs.LG 新提交 90%

From Tokens to States: LLMs as a Special Case of World Models and the Continuous Path Beyond

从令牌到状态:LLM作为世界模型的特例及其连续路径

Paul Dubois

机构 * Paul Dubois(保罗·杜博伊斯)

专题命中 其他LLM :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文论证LLM是世界模型的退化特例,并提出从NTP到JEPA的连续谱系,逐步放松LLM约束,同时探讨其可扩展性挑战。

Comments 10 pages, 6 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09303 2026-05-28 q-fin.PM 90%

Investor risk profiles of large language models

大型语言模型的投资者风险画像

Hanyong Cho, Geumil Bae, Jang Ho Kim

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 本研究通过标准化风险问卷评估GPT、Gemini和Llama三种大型语言模型的风险偏好,发现它们普遍为长期投资者但风险容忍度不同,且赋予特定人物角色后各模型会调整其风险画像。

Comments Poster presented at the AI for Finance Symposium '25, The 6th ACM International Conference on AI in Finance (ICAIF '25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15840 2026-04-28 cs.LO cs.FL 90%

Automatic Generation of Safety-compliant Linear Temporal Logic via Large Language Model: A Self-supervised Framework

基于大语言模型的自动安全合规线性时序逻辑生成:一种自监督框架

Junle Li, Siqi Chen, Jiakai Li, Meiqi Tian, Bingzhuo Zhong

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 本文提出AutoSafeLTL框架,利用大语言模型自动生成符合安全限制的LTL规范,通过语言包含检查与自动反例引导修改机制确保逻辑一致性和语义准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.04288 2026-04-16 cs.CR cs.SE 90%

LLM-Enabled Open-Source Systems in the Wild: An Empirical Study of Vulnerabilities in GitHub Security Advisories

基于大语言模型的开源系统在现实中的应用:对GitHub安全通告中漏洞的实证研究

Fariha Tanjim Shifat, Hariswar Baburaj, Ce Zhou, Jaydeb Sarker, Mia Mohammad Imran

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract,comments);language model(abstract,comments)

AI总结 本研究分析了295份GitHub安全通告,发现大多数漏洞映射到已知CWE,但模型中介暴露不足,建议结合CWE和OWASP视角更全面地评估LLM集成系统的漏洞。

Comments The 2nd International Workshop on Large Language Model Supply Chain Analysis (LLMSC 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14826 2026-04-07 cs.CL cs.AI cs.HC cs.LG 90%

SPRIG: Improving Large Language Model Performance by System Prompt Optimization

SPRIG:通过系统提示优化提升大语言模型性能

Lechen Zhang, Tolga Ergen, Lajanugen Logeswaran, Moontae Lee, David Jurgens

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) University of Michigan(密歇根大学) University of Illinois Chicago(伊利诺伊大学芝加哥分校) LG AI Research(LG AI 研究院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出SPRIG,一种基于编辑的遗传算法,通过优化通用提示提升大语言模型性能,发现系统提示与任务提示结合可进一步提升效果。

Comments Accepted at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15437 2026-01-23 cs.HC 90%

Exploring Implicit Perspectives on Autism in Large Language Models Through Multi-Agent Simulations

通过多智能体模拟探索大型语言模型中对自闭症的隐含视角

Sohyeon Park, Jesus Armando Beltran, Aehong Min, Anamara Ritt-Olson, Gillian R. Hayes

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

AI总结 通过多智能体模拟研究大型语言模型对自闭症的隐含视角,揭示其偏见并提出双重共情问题以改进与自闭症人群的互动。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20408 2026-01-08 q-bio.NC cs.AI cs.CL cs.LG 90%

Brain-Inspired Exploration of Functional Networks and Key Neurons in Large Language Models

受脑启发的大型语言模型中功能网络和关键神经元探索

Yiheng Liu, Zhengliang Liu, Zihao Wu, Junhao Ning, Haiyang Sun, Sichen Xia, Yang Yang, Xiaohui Gao, Ning Qiang, Bao Ge, Tianming Liu, Junwei Han, Xintao Hu

机构 * School of Automation, Northwestern Polytechnical University, Xi’an, China(自动化学院,西北工业大学,西安,中国) School of Computing, University of Georgia, Athens, USA(计算机学院,佐治亚大学,亚特兰大,美国) School of Physics and Information Technology, Shaanxi Normal University, Xi’an, China(物理与信息技术学院,陕西师范大学,西安,中国)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文受脑启发,探索LLM中的功能网络和关键神经元,发现这些网络对模型性能至关重要,通过抑制或增强网络活动可影响模型整体表现或特定任务效果。

Comments 21 pages, 18 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06192 2025-12-03 cs.DB cs.AI cs.CL cs.LG 90%

SQLBarber: A System Leveraging Large Language Models to Generate Customized and Realistic SQL Workloads

SQLBarber: 借助大型语言模型生成定制化和真实SQL工作负载的系统

Jiale Lao, Immanuel Trummer

机构 * Cornell University(康奈尔大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 SQLBarber利用大型语言模型生成定制化且真实的SQL工作负载,通过声明式接口和贝叶斯优化器高效生成符合目标成本分布的查询。

Comments Accepted by SIGMOD 2026; extended version with appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.14033 2025-11-18 cs.AI cs.CL cs.LG 90%

MLR-Copilot: Autonomous Machine Learning Research based on Large Language Models Agents

Ruochen Li, Teerth Patel, Qingyun Wang, Xinya Du

机构 * University of Texas at Dallas(德克萨斯大学达拉斯分校) UIUC(伊利诺伊大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12830 2025-08-15 cs.CL cs.AI cs.LG 90%

Knowledge-based Consistency Testing of Large Language Models

Sai Sathiesh Rajan, Ezekiel Soremekun, Sudipta Chattopadhyay

机构 * Singapore University of Technology and Design, Singapore(新加坡科技设计大学) Royal Holloway, University of London, UK(伦敦大学皇家霍洛威学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 12 pages, 4 figures, 8 tables, Accepted at EMNLP 2024 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12856 2025-08-12 cs.CL cs.AI cs.CY cs.LG 90%

AI-AI Bias: large language models favor communications generated by large language models

Walter Laurito, Benjamin Davis, Peli Grietzer, Tomáš Gavenčiak, Ada Böhm, Jan Kulveit

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 8 pages, 4 figures

Journal ref Proc. Natl. Acad. Sci. U.S.A. 122 (31) e2415697122 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11316 2025-07-16 cs.CL cs.AI cs.LG 90%

Internal Value Alignment in Large Language Models through Controlled Value Vector Activation

Haoran Jin, Meng Li, Xiting Wang, Zhihao Xu, Minlie Huang, Yantao Jia, Defu Lian

机构 * University of Science and Technology of China(中国科学技术大学) State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室) Gaoling School of Artificial Intelligence(光明人工智能学院) Renmin University of China(中国人民大学) Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大型模型与智能治理重点实验室) Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程技术研究中心) Tsinghua University(清华大学) Huawei Technologies Co. Ltd(华为技术有限公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 25 pages, 14 figures. Accepted by ACL 2025 (main conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.03247 2025-06-13 cs.HC 90%

End User Authoring of Personalized Content Classifiers: Comparing Example Labeling, Rule Writing, and LLM Prompting

Leijie Wang, Kathryn Yurechko, Pranati Dani, Quan Ze Chen, Amy X. Zhang

专题命中 其他LLM :LLM(title,abstract);prompting(title,abstract);large language model(abstract);language model(abstract)

Comments Accepted by CHI'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07176 2025-06-03 cs.CL cs.AI cs.LG 90%

Astute RAG: Overcoming Imperfect Retrieval Augmentation and Knowledge Conflicts for Large Language Models

Fei Wang, Xingchen Wan, Ruoxi Sun, Jiefeng Chen, Sercan Ö. Arık

机构 * Google Cloud AI Research(谷歌云人工智能研究) University of Southern California(南加州大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments ACL 2025 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09945 2025-05-16 cs.CL cs.AI cs.LG 90%

Personalizing Large Language Models using Retrieval Augmented Generation and Knowledge Graph

Deeksha Prahlad, Chanhee Lee, Dongha Kim, Hokeun Kim

机构 * Arizona State University(亚利桑那州立大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments To appear in the Companion Proceedings of the ACM Web Conference 2025 (WWW Companion '25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.16877 2025-04-24 cs.SE 90%

Context-Enhanced Vulnerability Detection Based on Large Language Model

Yixin Yang, Bowen Xu, Xiang Gao, Hailong Sun

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04520 2025-04-08 cs.LG cs.AI cs.CL 90%

Hessian of Perplexity for Large Language Models by PyTorch autograd (Open Source)

Ivan Ilin

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 15 pages, 3 figures, open source code on GitHub

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02055 2025-04-04 cs.DB 90%

MageSQL: Enhancing In-context Learning for Text-to-SQL Applications with Large Language Models

Chen Shen, Jin Wang, Sajjadur Rahman, Eser Kandogan

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19260 2025-03-26 cs.CL cs.AI cs.LG 90%

Linguistic Blind Spots of Large Language Models

Jiali Cheng, Hadi Amiri

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments NAACL 2025 Cognitive Modeling and Computational Linguistics Workshop

Journal ref NAACL 2025 CMCL Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12951 2025-03-18 cs.CV 90%

On the Consistency of Video Large Language Models in Temporal Comprehension

Minjoon Jung, Junbin Xiao, Byoung-Tak Zhang, Angela Yao

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

Comments Accepted to CVPR'25

详情

展开后加载摘要…

URL PDF HTML 收藏