arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12157 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12157 篇

2305.05576 2023-05-10 cs.CY cs.CL 88%

Large Language Models Humanize Technology

Pratyush Kumar

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.09337 2023-04-20 cs.HC cs.AI cs.MM 88%

Promptify: Text-to-Image Generation through Interactive Prompt Exploration with Large Language Models

Stephen Brade, Bryan Wang, Mauricio Sousa, Sageev Oore, Tovi Grossman

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.09181 2023-04-20 cs.SE cs.AI 88%

Large Language Models Based Automatic Synthesis of Software Specifications

Shantanu Mandal, Adhrik Chethan, Vahid Janfaza, S M Farabi Mahmud, Todd A Anderson, Javier Turek, Jesmin Jahan Tithi, Abdullah Muzahid

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.14828 2023-03-01 cs.CL 88%

Automatic Scoring of Dream Reports' Emotional Content with Large Language Models

Lorenzo Bertolini, Valentina Elce, Adriana Michalak, Giulio Bernardi, Julie Weeds

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.12834 2023-02-28 cs.HC cs.AI 88%

Leveraging Large Language Model and Story-Based Gamification in Intelligent Tutoring System to Scaffold Introductory Programming Courses: A Design-Based Research Study

Chen Cao

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

Comments Doctoral consortium for IUI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.10297 2023-01-26 cs.CL q-bio.NC 88%

Large language models can segment narrative events similarly to humans

Sebastian Michelmann, Manoj Kumar, Kenneth A. Norman, Mariya Toneva

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.09895 2022-12-21 cs.CL 88%

Improved Long-Form Spoken Language Translation with Large Language Models

Arya D. McCarthy, Hao Zhang, Shankar Kumar, Felix Stahlberg, Axel H. Ng

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.11445 2022-11-29 cs.CL 88%

Induced Natural Language Rationales and Interleaved Markup Tokens Enable Extrapolation in Large Language Models

Mirelle Bueno, Carlos Gemmell, Jeffrey Dalton, Roberto Lotufo, Rodrigo Nogueira

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.08536 2022-10-18 cs.CL 88%

Knowledge Prompting in Pre-trained Language Model for Natural Language Understanding

Jianing Wang, Wenkang Huang, Qiuhui Shi, Hongbin Wang, Minghui Qiu, Xiang Li, Ming Gao

专题命中 其他LLM :language model(title,abstract);prompting(title,abstract);分类 cs.CL

Comments 14 pages, 5 figures. This paper has been accepted for the main conference of EMNLP2022 (long paper)

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.06609 2022-03-18 cs.CL 88%

MSP: Multi-Stage Prompting for Making Pre-trained Language Models Better Translators

Zhixing Tan, Xiangwen Zhang, Shuo Wang, Yang Liu

专题命中 其他LLM :language model(title,abstract);prompting(title,abstract);分类 cs.CL

Comments ACL 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09981 2024-05-17 cs.CV 88%

Adversarial Robustness for Visual Grounding of Multimodal Large Language Models

Kuofeng Gao, Yang Bai, Jiawang Bai, Yong Yang, Shu-Tao Xia

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);foundation model(comments)

Comments ICLR 2024 Workshop on Reliable and Responsible Foundation Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.10480 2026-08-12 cs.AI cs.LG 新提交 88%

Multi-Granular Rationale-Guided Molecular LLM for Property Prediction

多粒度理由引导的分子大语言模型用于属性预测

Junwoo Park, Minyoung Shin, Cheol Soon Lee, Sujee Lee

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本研究提出多粒度理由引导的分子大语言模型MR-MoL,将GNN导出的子结构属性归因作为证据,在8项MoleculeNet任务上取得通用模型最佳结果,缩小了与专用模型的差距。

Comments 16 pages, 5 figures, 18 tables. Code: https://github.com/skku-aihclab/MR-MoL

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02985 2026-08-05 cs.LG cs.CL stat.ML 新提交 88%

Temporal Leakage in LLM Backtesting: Measurement, Validation, and Adjusted Scores

大语言模型回测中的时间泄漏:测量、验证与调整后得分

Zeyu Zhang, Bradly C. Stadie

机构 * Northwestern University(西北大学)

专题命中 其他LLM :LLM(title,summary_cn);分类 cs.CL、cs.LG

AI总结 该研究指出LLM回测的标准污染检测方法无效,提出利用已知截止时间和匹配干净对照组测量时间泄漏的方法,可检测并调整回测得分,澄清部分模型优势源于近期性而非真实技能。

Comments 12 pages main content, 45 pages in total

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.08054 2026-07-10 cs.LG cs.AI 新提交 88%

Who Analyses the Analyser? Self-Validating LLM Hazard Analysis with Constitutional Meta-STPA

谁来分析分析器?使用结构化元安全分析过程的自验证大语言模型风险分析

Samuel Tetteh, Udip Shrestha, Joshua R. Waite, Cody Fleming

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 研究针对大语言模型辅助安全分析工具自身缺乏分析的问题,提出结构化元安全分析过程,通过运行元安全分析过程推导治理章程,包含21条工具原则和8条元安全原则,还报告了自推导等四个发现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25852 2026-06-25 cs.LG cs.AI 新提交 88%

Semantic Consistency Policy Optimization for Reinforcement Learning of LLM Agents

语义一致性策略优化:用于LLM智能体强化学习

Peng Xu, Sijia Chen, Junzhuo Li, Xuming Hu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) The Hong Kong University of Science and Technology(香港科技大学)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.AI、cs.LG

AI总结 针对基于组的强化学习中语义一致的中间步骤因轨迹成败获得相反信用的问题,提出无价值奖励塑造方法SCPO,通过从组内成功兄弟步骤恢复信用,在ALFWorld和WebShop上达到93.7%和74.8%的成功率。

Comments 16 pages, 7 figures, 5 tables. Under review at EMNLP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22475 2026-06-23 cs.SE cs.AI cs.LG 新提交 88%

All Green, Still Broken: Real-Flow Verification Lessons from an LLM-Integrated, Multi-Market Web Application

全绿,依然破碎:来自一个集成LLM的多市场Web应用的实际流程验证教训

Muhammad Bilal, Ali Hassaan Mughal

机构 * Technical University of Munich(慕尼黑技术大学) Independent Researcher(独立研究员)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文通过一个生产级租赁搜索助手项目,分析了252个bug修复提交,发现44%的缺陷逃逸于组件级单元测试无法覆盖的四个接缝:实时浏览器运行时、非默认市场、端到端流程和全系统级别。提出了四接缝框架和缺陷分布测量,并分享了团队实践。

Comments 7 pages, 4 figures, 2 tables. Preprint of a manuscript submitted to IEEE Software

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.29440 2026-05-29 cs.CL cs.AI cs.IR 88%

SkillBrew: Multi-Objective Curation of Skill Banks for LLM Agents

SkillBrew: LLM智能体技能库的多目标策展

Wentao Hu, Zhendong Chu, Yiming Zhang, Junda Wu, Ming Jin, Xiangyu Zhao, Yilei Shao, Yanfeng Wang, Qingsong Wen

机构 * City University of Hong Kong(香港城市大学) Squirrel Ai Learning University of Science and Technology of China(中国科学技术大学) University of California, San Diego(加州大学圣地亚哥分校) Griffith University(格里菲斯大学) East China Normal University(华东师范大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.CL、cs.AI

AI总结 提出SkillBrew框架,将技能库策展建模为带效用约束的帕累托优化问题,通过双层提议-验证循环实现技能库的精简与多样性。

Comments 16 pages. Preprint. Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24490 2026-05-26 cs.AI cs.LG q-fin.PM 88%

Market Regime Council for Dynamic Credit Assignment in Multi-Agent LLM Decision Systems

市场制度委员会:多智能体LLM决策系统中的动态信用分配

Yunhua Pei, Zerui Ge, Jin Zheng, John Cartlidge

机构 * University of Bristol, UK(布里斯托大学)

专题命中 其他LLM :LLM(title,title_cn);分类 cs.AI、cs.LG

AI总结 提出市场制度委员会(MRC),一种基于Shapley值进行在线智能体加权、贝叶斯自适应混合和制度依赖乘数的多智能体决策系统,在加密货币投资中实现高夏普比率和累计收益。

Comments 35 pages, 13 figures, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24048 2026-05-26 cs.LG cs.AI 88%

Mixture of Complementary Agents for Robust LLM Ensemble

互补代理混合:鲁棒的大语言模型集成

Yichi Zhang, Kevin Lu, Yuang Zhang, Jie Gao, Lirong Xia, Fang-Yi Yu

机构 * DIMACS, Rutgers University(罗格斯大学DIMACS研究中心) Department of Mathematics, Rutgers University(罗格斯大学数学系) Department of Computer Science, George Mason University(乔治·梅森大学计算机科学系) Department of Computer Science, Rutgers University(罗格斯大学计算机科学系)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 将大语言模型选择视为组合选择问题,提出基于互补性的贪心选择算法,在性能与成本间取得最佳平衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07235 2026-05-25 cs.LG cs.AI cs.IT math.IT 88%

ArcMark: Distortion-Free Multi-Byte LLM Watermark via Optimal Transport

ArcMark: 通过最优传输实现无失真的多字节大语言模型水印

Atefeh Gilani, Sajani Vithana, Carol Xuan Long, Oliver Kosut, Lalitha Sankar, Flavio P. Calmon

机构 * Arizona State University(亚利桑那州立大学) Harvard University(哈佛大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出ArcMark,一种基于编码和信息论原理的无失真水印方法,能够在数百个token中可靠嵌入多字节信息,且不改变LLM的下一token分布。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00087 2026-05-04 cs.NI cs.AI cs.CY cs.IR cs.LG 88%

DeGenTWeb: A First Look at LLM-dominant Websites

DeGenTWeb:对以大语言模型为主导的网站的首次观察

Sichang Steven He, Calvin Ardi, Ramesh Govindan, Harsha V. Madhyastha

机构 * University of Southern California(南加州大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文通过DeGenTWeb系统识别出大量由大语言模型生成内容的网站,发现其在Web中的普及率持续增长,但准确识别此类网站面临挑战。

Comments 6 pages, 6 figures, 13 page total; in submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06587 2026-04-29 cs.CL cs.AI cs.HC 88%

TouchAI: Exploring human-AI perceptual alignment in touch through language model representations

TouchAI: 探索通过语言模型表示的触觉人机感知对齐

Shu Zhong, Elia Gatti, Youngjun Cho, Marianna Obrist

机构 * Department of Computer Science, University College London(伦敦大学学院计算机科学系)

专题命中 其他LLM :language model(title,abstract);LLM(abstract,abstract_cn);large language model(abstract);分类 cs.CL、cs.AI

AI总结 本文通过'textile hand'任务研究大型语言模型在触觉感知对齐中的表现,发现不同织物样本的对齐程度差异显著,揭示了人机感知对齐的挑战与潜力。

Comments Accepted at IJHCS

Journal ref International Journal of Human-Computer Studies 210 (2026) 103765

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13829 2026-04-20 cs.CL cs.AI 88%

A Linguistics-Aware LLM Watermarking via Syntactic Predictability

一种基于语言学的LLM水印方法:通过语法可预测性

Shinwoo Park, Hyejin Park, Hyeseon An, Yo-Sub Han

机构 * Yonsei University(延世大学) Rensselaer Polytechnic Institute(拉特格斯理工学院)

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出STELA框架,通过语法可预测性动态调整水印信号,提升检测鲁棒性,适用于多种语言。

Comments ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.16848 2026-03-10 cs.LG cs.AI 88%

Meta-RL Induces Exploration in Language Agents

元强化学习诱导语言智能体的探索

Yulun Jiang, Liangze Jiang, Damien Teney, Michael Moor, Maria Brbic

机构 * EPFL(苏黎世联邦理工学院) ETH Zurich(苏黎世联邦理工学院) Idiap Research Institute(Idiap研究机构)

专题命中 其他LLM :language agent(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

AI总结 LaMer通过元强化学习框架提升语言智能体的探索能力,显著提高多任务性能并增强泛化能力。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13499 2026-02-09 cs.HC cs.AI cs.LG 88%

Deception at Scale: Deceptive Designs in 1K LLM-Generated Ecommerce Components

大规模欺骗:1000个LLM生成的电商组件中的欺骗设计

Ziwei Chen, Jiawen Shen, Luna, Hanyu Zhang, Kristen Vaccaro

机构 * University of California San Diego(加州大学圣地亚哥分校)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 研究揭示LLM生成电商组件中普遍存在欺骗性设计,通过测试不同提示策略发现价值观导向方法最有效,强调LLM在编码中的潜在风险及改进方向。

Comments 18 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13302 2025-12-19 cs.CL cs.AI 88%

LLM one-shot style transfer for Authorship Attribution and Verification

LLM单次风格迁移用于作者身份识别与验证

Pablo Miralles-González, Javier Huertas-Tato, Alejandro Martín, David Camacho

机构 * Department of Computer Systems(计算机系统系) Technical University of Madrid(马德里技术大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文提出利用LLM的单次风格迁移能力,通过无监督方法提升作者身份识别与验证的性能,同时优化计算成本与任务表现的平衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06991 2025-12-09 cs.CL cs.AI 88%

Prompting-in-a-Series: Psychology-Informed Contents and Embeddings for Personality Recognition With Decoder-Only Models

Prompting-in-a-Series: 基于心理启发内容与嵌入的个性识别解码器-only模型

Jing Jie Tan, Ban-Hoe Kwan, Danny Wee-Kiat Ng, Yan-Chai Hum, Anissa Mokraoui, Shih-Yu Lo

机构 * Laboratoire de traitement et transport de l'information, Université Sorbonne Paris Nord, France(信息处理与传输实验室,索邦巴黎北大学,法国) Institute of Communication Studies, National Yang Ming Chiao Tung University, Taiwan(传播学研究所,国立阳明交通大学,台湾)

专题命中 其他LLM :prompting(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

AI总结 PICEPR通过心理启发内容与嵌入技术,提升解码器-only模型在个性识别中的性能,实现5-15%的提升。

Comments 16 pages

Journal ref IEEE Transactions on Computational Social Systems, pages 1-15, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13143 2025-10-16 cs.CL cs.AI 88%

Stable LLM Ensemble: Interaction between Example Representativeness and Diversity

Junichiro Niimi

机构 * Meijo University(名古屋大学) RIKEN AIP(理化学研究所AIP)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02425 2025-10-06 cs.CL cs.CV cs.LG 88%

Words That Make Language Models Perceive

Sophie L. Wang, Phillip Isola, Brian Cheung

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20258 2025-09-16 cs.CL cs.AI 88%

LLM as a Broken Telephone: Iterative Generation Distorts Information

Amr Mohamed, Mingmeng Geng, Michalis Vazirgiannis, Guokan Shang

机构 * MBZUAI SISSA Ecole Polytechnique(巴黎政治学院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Accepted to ACL 2025, Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏