arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12157 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12157 篇

2602.06056 2026-02-09 cs.MM cs.AI cs.CL cs.CV 86%

Analyzing Diffusion and Autoregressive Vision Language Models in Multimodal Embedding Space

分析扩散模型和自回归视觉语言模型在多模态嵌入空间中的表现

Zihang Wang, Siyue Zhang, Yilun Zhao, Jingyi Yang, Tingyu Song, Anh Tuan Luu, Chen Zhao

机构 * Nanyang Technological University(南洋理工大学) Yale University(耶鲁大学) NYU Shanghai(纽约大学上海分校) Alibaba-NTU Singapore Joint Research Institute(阿里-国立新加坡大学联合研究机构) University of the Chinese Academy of Sciences(中国科学院大学) Center for Data Science(数据科学中心) New York University(纽约大学)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);foundation model(abstract);分类 cs.CL、cs.AI

AI总结 本文研究了多模态扩散模型在多模态嵌入任务中的表现,发现其在分类、VQA和检索任务中均逊于自回归VLM。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03969 2026-02-06 cs.AI cs.CR cs.LG 86%

How Catastrophic is Your LLM? Certifying Risk in Conversation

你的LLM有多危险?对话中风险的认证

Chengxiao Wang, Isha Chaudhary, Qian Hu, Weitong Ruan, Rahul Gupta, Gagandeep Singh

机构 * University of Illinois, Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本研究提出C$^3$LLM框架,通过统计方法认证LLMs在多轮对话中的灾难性风险,揭示前沿模型中高达70%的潜在风险,强调改进安全训练的必要性。

Comments Accepted by ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23242 2026-02-05 cs.CL cs.AI 86%

Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation

超越猜测:测量大型语言模型生成文本在多语言虚假信息中的增长

Dominik Macko, Aashish Anantha Ramakrishnan, Jason Samuel Lucas, Robert Moro, Ivan Srba, Adaku Uchendu, Dongwon Lee

机构 * Kempelen Institute of Intelligent Technologies(智能技术研究所) The Pennsylvania State University(宾夕法尼亚州立大学) MIT Lincoln Laboratory(麻省理工学院林赛实验室)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究通过实证数据证明了LLM在多语言虚假信息中的增长,揭示了不同语言、平台和时间周期中的关键模式。

Comments accepted to Computer magazine

Journal ref Computer (Volume: 59, Issue: 2, February 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01063 2026-02-03 cs.CL cs.AI 86%

Personality Expression Across Contexts: Linguistic and Behavioral Variation in LLM Agents

人格表达的多情境性:语言与行为在大语言模型代理中的变化

Bin Han, Deuksin Kwon, Jonathan Gratch

机构 * University of Southern California(南加州大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究探讨了大语言模型在不同对话情境中人格表达的差异,发现相同人格提示在不同情境下产生不同语言、行为和情感表现,表明LLMs能根据社交需求灵活调整。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01025 2026-02-03 cs.LG cs.AI cs.CV 86%

Toward Universal and Transferable Jailbreak Attacks on Vision-Language Models

迈向通用且可迁移的视觉-语言模型劫持攻击

Kaiyuan Cui, Yige Li, Yutao Wu, Xingjun Ma, Sarah Erfani, Christopher Leckie, Hanxun Huang

机构 * School of Computing and Information Systems, The University of Melbourne, Australia(墨尔本大学计算机与信息系) School of Computing and Information Systems, Singapore Management University, Singapore(新加坡管理大学计算机与信息系) School of Information Technology, Deakin University, Australia(德肯大学信息科技系) Institute of Trustworthy Embodied AI, Fudan University, China(复旦大学可信具身人工智能研究所)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出UltraBreak,一种通用且可迁移的视觉-语言模型劫持攻击框架,通过视觉正则化和语义引导的文本监督,有效提升攻击的泛化能力和转移性能。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00282 2026-01-29 cs.AI cs.CL 86%

Mind the Gap: The Divergence Between Human and LLM-Generated Tasks

注意差距:人类与LLM生成任务之间的差异

Yi-Long Lu, Jiajun Song, Chunhui Zhang, Wei Wang

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 研究揭示了人类与LLM在任务生成中的核心差异,指出人类受心理驱动影响,而LLM在生成具身目标方面存在不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.16087 2026-01-23 cs.AI cs.CL 86%

Controlling Long-Horizon Behavior in Language Model Agents with Explicit State Dynamics

通过显式状态动力学控制语言模型代理的长周期行为

Sukesh Subaharan

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出通过显式情感动力学控制语言模型代理的长周期行为,通过引入持续的情感状态并结合一阶和二阶更新规则,提升对话的连贯性和响应稳定性。

Comments Supplementary materials can be found here: https://github.com/drsukeshs/agent-behavior-ext-dynamics

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13855 2026-01-21 cs.CL cs.AI 86%

Harnessing Consistency for Robust Test-Time LLM Ensemble

利用一致性提升鲁棒性测试时LLM集成

Zhichen Zeng, Qi Yu, Xiao Lin, Ruizhong Qiu, Xuying Ning, Tianxin Wei, Yuchen Yan, Jingrui He, Hanghang Tong

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 CoRE通过利用模型一致性提升LLM集成的鲁棒性,通过token和model级别的一致性改进集成性能。

Comments 18 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11572 2026-01-21 cs.LG cs.AI 86%

Discrete Semantic States and Hamiltonian Dynamics in LLM Embedding Spaces

离散语义状态与哈密顿动力学在大语言模型嵌入空间中的应用

Timo Aukusti Laine

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本研究通过哈密顿动力学分析LLM嵌入空间的结构,揭示了语义状态的离散性及潜在的量子力学联系。

Comments 23 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11342 2026-01-19 cs.LG cs.CL 86%

Unlocking the Potentials of Retrieval-Augmented Generation for Diffusion Language Models

解锁检索增强生成在扩散语言模型中的潜力

Chuanyue Yu, Jiahui Wang, Yuhan Li, Heng Chang, Ge Lan, Qingyun Sun, Jia Li, Jianxin Li, Ziwei Zhang

机构 * Nankai University(南开大学) Beihang University(北航) HKUST (Guangzhou)(香港科技大学(广州)) Huawei Technologies Co., Ltd.(华为技术有限公司)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出SPREAD框架,通过引入查询相关性引导的去噪策略,解决DLMs在RAG框架中生成精度低和语义漂移的问题。

Comments Preprints

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.03747 2026-01-08 cs.LG cs.CL stat.AP 86%

Context-Alignment: Activating and Enhancing LLM Capabilities in Time Series

上下文对齐:在时间序列中激活和增强大语言模型的能力

Yuxiao Hu, Qian Li, Dongxiao Zhang, Jinyue Yan, Yuntian Chen

机构 * The Hong Kong Polytechnic University(香港理工大学) Ningbo Institute of Digital Twin(宁波数字孪生研究所) Eastern Institute of Technology(东部技术研究所) Shanghai Jiao Tong University(上海交通大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文提出上下文对齐方法,通过多模态输入和图神经网络增强LLMs在时间序列任务中的能力,提升逻辑和结构理解,提高预测性能。

Comments This paper has been accepted by ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.05499 2025-12-30 cs.CR cs.AI cs.CL cs.SE 86%

Prompt Injection attack against LLM-integrated Applications

针对集成大语言模型应用的提示注入攻击

Yi Liu, Gelei Deng, Yuekang Li, Kailong Wang, Zihao Wang, Xiaofeng Wang, Tianwei Zhang, Yepang Liu, Haoyu Wang, Yan Zheng, Leo Yu Zhang, Yang Liu

机构 * Griffith University(格里菲斯大学) Nanyang Technological University(南洋理工大学) University of New South Wales(新南威尔士大学) Huazhong University of Science and Technology(华中科技大学) Indiana University at Bloomington(印第安纳大学布卢明顿分校) Southern University of Science and Technology(南方科技大学) Tianjin University(天津大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究提出HouYi技术,揭示LLM集成应用中提示注入攻击的潜在风险及缓解方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07462 2025-12-15 cs.MA cs.AI cs.GT cs.LG math.DS 86%

Understanding LLM Agent Behaviours via Game Theory: Strategy Recognition, Biases and Multi-Agent Dynamics

通过博弈论理解LLM代理行为:策略识别、偏差与多代理动态

Trung-Kiet Huynh, Duy-Minh Dao-Sy, Thanh-Bang Cao, Phong-Hao Le, Hong-Dan Nguyen, Phu-Quy Nguyen-Lam, Minh-Luan Nguyen-Vo, Hong-Phat Pham, Phu-Hoa Pham, Thien-Kim Than, Chi-Nguyen Tran, Huy Tran, Gia-Thoai Tran-Le, Alessio Buscemi, Le Hong Trang, The Anh Han

机构 * Faculty of Information and Technology, Ho Chi Minh City University of Science (HCMUS), Vietnam(信息科技学院,胡志明市科学大学(HCMUS)) Vietnam National University - Ho Chi Minh City (VNU-HCM), Vietnam(越南国家大学-胡志明市(VNU-HCM)) Faculty of Computer Science and Engineering, Ho Chi Minh City University of Technology (HCMUT), Vietnam(计算机科学与工程学院,胡志明市技术大学(HCMUT))

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文通过扩展FAIRGAME框架,系统评估LLM在重复社会困境中的行为,揭示其策略识别、偏差及多代理动态,为AI治理和安全多代理系统设计提供方法论基础。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05331 2025-12-08 cs.CL cs.LG 86%

Exposing Pink Slime Journalism: Linguistic Signatures and Robust Detection Against LLM-Generated Threats

揭露粉红 slime 纪实:语言特征与对抗 LLM 生成威胁的鲁棒检测

Sadat Shahriar, Navid Ayoobi, Arjun Mukherjee, Mostafa Musharrat, Sai Vishnu Vamsi

机构 * University of Houston, Texas, USA(德克萨斯大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出了一种基于语言特征的鲁棒检测框架,以应对LLM生成的粉红 slime 纪实威胁,提升了检测性能27%。

Comments Published in RANLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00293 2025-12-02 cs.LG cs.AI 86%

FiCoTS: Fine-to-Coarse LLM-Enhanced Hierarchical Cross-Modality Interaction for Time Series Forecasting

FiCoTS: 细到粗的LLM增强层次跨模态交互用于时间序列预测

Yafei Lyu, Hao Zhou, Lu Zhang, Xu Yang, Zhiyong Liu

机构 * School of Advanced Interdisciplinary Sciences, University of Chinese Academy Sciences(中国科学院大学先进交叉学科学院) MAIS, Institute of Automation, Chinese Academy of Science(中国科学院自动化研究所MAIS) Great Bay University(大亚大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 FiCoTS通过细到粗的LLM增强层次跨模态交互框架,提升多模态时间序列预测的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21066 2025-11-27 cs.CL cs.AI 86%

Context-Aware Pragmatic Metacognitive Prompting for Sarcasm Detection

面向上下文的元认知修辞提示用于讽刺检测

Michael Iskandardinata, William Christian, Derwin Suhartono

机构 * Computer Science Department(计算机科学系) School of Computer Science(计算机科学学院) Bina Nusantara University(宾努斯大学)

专题命中 其他LLM :prompting(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出一种基于检索的元认知修辞提示方法,通过整合上下文信息提升LLMs在讽刺检测中的性能,实验显示在多个数据集上均取得显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18038 2025-11-25 cs.SE cs.AI cs.LG 86%

MASTEST: A LLM-Based Multi-Agent System For RESTful API Tests

MASTEST:基于LLM的多智能体系统用于RESTful API测试

Xiaoke Han, Hong Zhu

机构 * School of Engineering, Computing and Mathematics, Oxford Brookes University(工程、计算与数学学院,奥克斯伯里大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 MASTEST利用LLM和编程智能体构建多智能体系统,实现RESTful API测试的全流程自动化,通过生成测试场景、脚本及分析响应,验证LLM在测试任务中的高效性与准确性。

Comments 14 Page of main text plus 4 pages of appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17671 2025-11-25 cs.CR cs.AI cs.CL 86%

MURMUR: Using cross-user chatter to break collaborative language agents in groups

利用跨用户交流打破协作语言代理组

Atharv Singh Patlan, Peiyao Sheng, S. Ashwin Hebbar, Prateek Mittal, Pramod Viswanath

机构 * Princeton University(普林斯顿大学) Sentient

专题命中 其他LLM :language agent(title,abstract);LLM(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 MURMUR通过生成真实用户交互,揭示了跨用户污染攻击对多用户语言代理的威胁,并提出基于任务的聚类作为初步防御措施。

Comments 20 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11829 2025-11-18 cs.CL cs.AI cs.FL cs.LO 86%

Towards Autoformalization of LLM-generated Outputs for Requirement Verification

Mihir Gupte, Ramesh S

机构 * General Motors(通用汽车公司)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments To be submitted for publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11258 2025-11-17 cs.CL cs.AI 86%

KGQuest: Template-Driven QA Generation from Knowledge Graphs with LLM-Based Refinement

Sania Nayab, Marco Simoni, Giulio Rossolini, Andrea Saracino

机构 * Scuola Superiore Sant’Anna, Pisa, Italy(圣安娜高等学院) Sapienza University of Rome, Rome, Italy(罗马萨皮恩扎大学) Institute of Informatics and Telematics, National Research Council of Italy (CNR)(意大利信息与电信研究所,国家研究理事会(CNR))

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09719 2025-11-17 cs.LG cs.AI 86%

ICL-Router: In-Context Learned Model Representations for LLM Routing

Chenxu Wang, Hao Li, Yiqun Zhang, Linyao Chen, Jianhao Chen, Ping Jian, Peng Ye, Qiaosheng Zhang, Shuyue Hu

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08143 2025-11-12 cs.CL cs.AI 86%

Relation as a Prior: A Novel Paradigm for LLM-based Document-level Relation Extraction

Qiankun Pi, Yepeng Sun, Jicang Lu, Qinlong Fan, Ningbo Huang, Shiyu Wang

机构 * Information Engineering University(信息工程大学) Academy of Military Science(军事科学院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 17 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14205 2025-10-30 cs.CL cs.AI 86%

DPRF: A Generalizable Dynamic Persona Refinement Framework for Optimizing Behavior Alignment Between Personalized LLM Role-Playing Agents and Humans

Bingsheng Yao, Bo Sun, Yuanzhe Dong, Yuxuan Lu, Dakuo Wang

机构 * Northeastern University(东北大学) Stanford University(斯坦福大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments In Submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14571 2025-10-29 cs.CR cs.AI cs.CL cs.HC 86%

Learned, Lagged, LLM-splained: LLM Responses to End User Security Questions

Vijay Prakash, Kevin Lee, Arkaprabha Bhattacharya, Danny Yuxing Huang, Jessica Staddon

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 17 pages, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20039 2025-10-24 cs.HC cs.AI cs.CL cs.CY 86%

Beyond One-Way Influence: Bidirectional Opinion Dynamics in Multi-Turn Human-LLM Interactions

Yuyang Jiang, Longjie Guo, Yuchen Wu, Aylin Caliskan, Tanu Mitra, Hua Shen

机构 * University of Chicago(芝加哥大学) New York University(纽约大学) University of Washington(华盛顿大学) New York University Shanghai(纽约大学上海)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 26 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11704 2025-10-21 cs.CL cs.AI 86%

Adapting Chat Language Models Using Only Target Unlabeled Language Data

Atsuki Yamaguchi, Terufumi Morishita, Aline Villavicencio, Nikolaos Aletras

机构 * University of Sheffield(谢菲尔德大学) Hitachi, Ltd.(日立株式会社) University of Exeter(埃克塞特大学)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to TMLR

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07141 2025-10-17 cs.CL cs.AI 86%

Comparing Human and Language Models Sentence Processing Difficulties on Complex Structures

Samuel Joseph Amouyal, Aya Meltzer-Asscher, Jonathan Berant

机构 * Blavatnik School of Computer Science, Tel Aviv University, Israel(巴尔-艾塔夫大学计算机科学学院) Department of Linguistics, Tel Aviv University, Israel(巴尔-艾塔夫大学语言学系) Sagol School of Neuroscience, Tel Aviv University, Israel(巴尔-艾塔夫大学神经科学学院)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments Data and code will be released soon

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12367 2025-10-15 cs.CL cs.AI 86%

LLM-REVal: Can We Trust LLM Reviewers Yet?

Rui Li, Jia-Chen Gu, Po-Nien Kung, Heming Xia, Junfeng liu, Xiangwen Kong, Zhifang Sui, Nanyun Peng

机构 * Peking University(北京大学) University of California, Los Angeles(加州大学洛杉矶分校) The Hong Kong Polytechnic University(香港理工大学) StepFun AI

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01644 2025-10-13 cs.CL cs.AI cs.CY 86%

Machine Learning for Detection and Analysis of Novel LLM Jailbreaks

John Hawkins, Aditya Pramar, Rodney Beard, Rohitash Chandra

机构 * Centre for Artificial Intelligence and Innovation(人工智能与创新中心) Pingla Institute(平拉研究所) Transitional Artificial Intelligence Research Group(过渡人工智能研究组) UNSW(新南威尔士大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11997 2025-10-03 cs.LG cs.AI 86%

Can LLMs Find Fraudsters? Multi-level LLM Enhanced Graph Fraud Detection

Tairan Huang, Yili Wang, Qiutong Li, Changlong He, Jianliang Gao

机构 * Central South University(中南大学) Hongkong University of Science and Technology (Guangzhou)(香港科技大学(广州))

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments Accepted by ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏