arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12169 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12169 篇

2407.16891 2024-07-25 cs.CY cs.CL 81%

Cultural Value Differences of LLMs: Prompt, Language, and Model Size

Qishuai Zhong, Yike Yun, Aixin Sun

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.00808 2024-07-02 eess.SY cs.AI cs.SY 81%

Exploring a Physics-Informed Decision Transformer for Distribution System Restoration: Methodology and Performance Analysis

Hong Zhao, Jin Wei-Kocsis, Adel Heidari Akhijahani, Karen L Butler-Purry

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.13905 2024-06-21 cs.CL 81%

Persuasiveness of Generated Free-Text Rationales in Subjective Decisions: A Case Study on Pairwise Argument Ranking

Mohamed Elaraby, Diane Litman, Xiang Lorraine Li, Ahmed Magooda

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09839 2024-06-17 cs.CL cs.HC 81%

Rapport-Driven Virtual Agent: Rapport Building Dialogue Strategy for Improving User Experience at First Meeting

Muhammad Yeza Baihaqi, Angel García Contreras, Seiya Kawano, Koichiro Yoshino

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments will be presented at INTERSPEECH 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.03963 2024-06-07 cs.CL 81%

A + B: A General Generator-Reader Framework for Optimizing LLMs to Unleash Synergy Potential

Wei Tang, Yixin Cao, Jiahao Ying, Bo Wang, Yuyue Zhao, Yong Liao, Pengyuan Zhou

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);foundation model(abstract)

Comments Accepted to ACL'24 (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.11941 2024-06-04 cs.CL 81%

CoCo-Agent: A Comprehensive Cognitive MLLM Agent for Smartphone GUI Automation

Xinbei Ma, Zhuosheng Zhang, Hai Zhao

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);language agent(abstract)

Comments ACL'2024 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.02883 2023-11-07 cs.CL 81%

SQLPrompt: In-Context Text-to-SQL with Minimal Labeled Data

Ruoxi Sun, Sercan Ö. Arik, Rajarishi Sinha, Hootan Nakhost, Hanjun Dai, Pengcheng Yin, Tomas Pfister

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.18287 2023-10-24 cs.CV cs.CL 81%

LaFTer: Label-Free Tuning of Zero-shot Classifier using Language and Unlabeled Image Collections

M. Jehanzeb Mirza, Leonid Karlinsky, Wei Lin, Mateusz Kozinski, Horst Possegger, Rogerio Feris, Horst Bischof

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments NeurIPS 2023 (Camera Ready) - Project Page: https://jmiemirza.github.io/LaFTer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.17421 2023-10-12 cs.CV cs.CL 81%

The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision)

Zhengyuan Yang, Linjie Li, Kevin Lin, Jianfeng Wang, Chung-Ching Lin, Zicheng Liu, Lijuan Wang

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.11755 2022-09-26 cs.CL cs.IR 81%

Promptagator: Few-shot Dense Retrieval From 8 Examples

Zhuyun Dai, Vincent Y. Zhao, Ji Ma, Yi Luan, Jianmo Ni, Jing Lu, Anton Bakalov, Kelvin Guu, Keith B. Hall, Ming-Wei Chang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.03711 2025-02-07 cs.CL cs.AI cs.LG 81%

MultiQ&A: An Analysis in Measuring Robustness via Automated Crowdsourcing of Question Perturbations and Answers

Nicole Cho, William Watson

专题命中 其他LLM :LLM(abstract,comments);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments AAAI 2025 Workshop on Preventing and Detecting LLM Misinformation (PDLM) (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.26045 2026-08-04 cs.CL cs.AI 版本更新 81%

Confidence and Calibration of Activation Oracles for Reliable Interpretation of Language Model Internals

激活预言机的置信度与校准:用于语言模型内部的可信解释

Federico Torrielli, Peter Schneider-Kamp, Lukas Galke Poech

机构 * University of Turin(都灵大学) University of Southern Denmark(南丹麦大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文研究了6种激活预言机置信度估计方法,发现bootstrap模式频率在校准上优于其他方法(ECE 5.7% vs 25.5%),而log-prob基线可作为快速分诊信号。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24917 2026-08-03 cs.CL cs.LG 版本更新 81%

Estimating near-verbatim extraction risk in language models with decoding-constrained beam search

通过解码约束束搜索估计语言模型中的近原文提取风险

A. Feder Cooper, Mark A. Lemley, Christopher De Sa, Lea Duesterwald, Allison Casasola, Jamie Hayes, Katherine Lee, Daniel E. Ho, Percy Liang

机构 * AVERI Stanford(斯坦福大学) Cornell(康奈尔大学) Google DeepMind(谷歌DeepMind)

专题命中 其他LLM :language model(title);LLM(abstract_cn);分类 cs.CL、cs.LG

AI总结 本文提出解码约束束搜索方法,以有效估计语言模型中的近原文提取风险,揭示更多可提取序列及模型规模对风险的影响。

Comments COLM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.15713 2026-07-20 eess.SP cs.AI cs.LG 新提交 81%

Map as a Prompt: Learning Multi-Modal Spatial-Signal Foundation Models for Cross-scenario Wireless Localization

地图作为提示:学习用于跨场景无线定位的多模态空间信号基础模型

Yong Chu, Xun Zhou, Zenglin Xu, Hui Wang, Yue Yu

机构 * Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Pengcheng Laboratory(鹏城实验室) Shanghai Academy of AI for Science(上海人工智能科学研究院) Artificial Intelligence Innovation and Incubation Institute, Fudan University(复旦大学人工智能创新与孵化院)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 针对无线定位在不同环境中面临的挑战及现有方法的局限,提出多模态基础模型SigMap,通过循环自适应掩码策略和“地图即提示”框架学习无线表示并实现跨场景适应,实验证明其性能领先且零样本泛化能力强。

Comments 17pages, 9 figures, poster in International Conference on Learning Representations (ICLR), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.14113 2026-07-17 cs.CL cs.AI 新提交 81%

T5-CSBoost: Adversarial Perturbation Resistant LLM Fingerprinting

T5-CSBoost:抗对抗扰动的语言模型指纹识别

Gayan K. Kulatilleke, Mahsa Baktashmotlagh, Siamak Layeghy, Marius Portmann

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL、cs.AI

AI总结 研究针对AIGT检测器在多种干扰下准确性下降的问题,提出T5-CSBoost,通过引入辅助损失鼓励学习抗扰动风格表示,在多基准测试中达先进水平,对高强度对抗扰动鲁棒性增强,证明对比学习规范风格嵌入可构建更强大指纹识别系统。

Comments 14 pages, 3 datasets, code and data will be provided

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.14106 2026-07-17 cs.CL cs.AI 新提交 81%

Token Time Continuous Diffusion for Language Modeling

用于语言建模的令牌时间连续扩散

Parikshit Bansal, Sujay Sanghavi

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 研究提出令牌时间连续扩散(TTCD)语言模型,其在连续空间运行,引入令牌时间概念。该模型避免多令牌并行采样,能更好建模条件生成。实验表明,在高速加速时TTCD优于离散模型,在无条件和条件生成上有优势,数独求解也有类似成果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.07375 2026-07-09 cs.LG cs.AI 新提交 81%

On Adversarial Vulnerability of Vision-Language Models through the Lens of Intermediate Spectral Subspaces

基于中间谱子空间视角的视觉语言模型对抗脆弱性研究

Chethan Krishnamurthy Ramanaik, Tobias Callies, Michael Hecht, Eirini Ntoutsi

机构 * University of the Bundeswehr Munich(慕尼黑联邦国防军大学) University of Wrocław(弗罗茨瓦夫大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI、cs.LG

AI总结 研究基于Transformer的视觉语言模型对抗脆弱性,提出白盒谱子空间引导攻击(SSGRA),通过将中间表示与特定子空间对齐来攻击,实验显示其比现有基线更有效,还为模型对抗脆弱性提供谱解释,助力提升鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00852 2026-07-02 cs.CL cs.AI 新提交 81%

Recovering Input Text from Hidden States: Study of Gradient-Based Inversion of Decoder-Only Language Models

从隐藏状态恢复输入文本:基于梯度的解码器专用语言模型反演研究

Mikołaj Słowikowski, Maciej Witold Majewski

机构 * AGH University of Krakow, Faculty of Physics and Applied Computer Science(AGH克拉科夫大学物理与应用计算机科学学院)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 研究解码器专用语言模型最后层隐藏状态的反演问题,提出连续嵌入空间优化方法,通过离散损失评估恢复正确性,发现高频功能词是主要失败原因,内容词几乎完美恢复。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.19831 2026-06-19 cs.CL cs.LG 新提交 81%

Leverage Is Not Reach: A Control-Window Law for Single-Neuron Steering in Language Models

杠杆不等于可达性:语言模型中单神经元操控的控制窗口定律

Hongliang Liu

机构 * Palo Alto Networks

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

AI总结 提出预算归一化控制窗口框架,通过残差范数与写入范数之比定义的相干预算,预测单神经元干预何时产生连贯行为控制,并在15个神经元上验证了预测精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05316 2026-06-08 cs.LG cs.AI 版本更新 81%

Standard vs. Modular Sampling: Best Practices for Reliable LLM Unlearning

标准采样与模块化采样:可靠的大语言模型遗忘的最佳实践

Praveen Bushipaka, Lucia Passaro, Tommaso Cucinotta

机构 * Scuola Superiore Sant’Anna(圣安纳高等学院) University of Pisa(比萨大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI、cs.LG

AI总结 针对大语言模型遗忘中采样策略的不足,提出模块化实体级遗忘(MELU)策略,通过多样化邻居集和模块化采样平衡遗忘效果与模型效用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.04459 2026-06-04 cs.CR cs.AI cs.CC cs.CL 81%

Token Rankings are Unforgeable Language Model Signatures

Token排名是不可伪造的语言模型签名

Matthew Finlayson, Andreas Grivas, Xiang Ren, Swabha Swayamdipta

机构 * University of Southern California(南加州大学) University of Edinburgh(爱丁堡大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文发现语言模型的token排名(按概率排序)构成唯一且不可伪造的签名,并研究了在限制API下如何平衡签名展示与参数泄露。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.03780 2026-06-03 cs.CL cs.LG 81%

Expert-Aware Causal Tracing of Factual Recall in Sparse MoE Language Models

专家感知的稀疏MoE语言模型中事实回忆的因果追踪

Yuetian Lu, Ali Modarressi, Yihong Liu, Hinrich Schütze

机构 * Center for Information and Language Processing (CIS)(信息与语言处理中心) Ubiquitous Knowledge Processing Lab (UKP)(无所不在的知识处理实验室) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

AI总结 针对稀疏混合专家语言模型,提出专家感知的因果追踪方法,通过干预专家级更新定位事实回忆的关键专家,发现专家级定位依赖于模型和协议。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00283 2026-05-29 cs.LG cs.AI q-bio.QM 81%

BioArc: Discovering Optimal Neural Architectures for Biological Foundation Models

BioArc:发现生物学基础模型的最优神经架构

Yi Fang, Haoran Xu, Jiaxin Han, Sirui Ding, Yizhi Wang, Yue Wang, Xuan Wang

机构 * Department of Computer Science, Virginia Tech, Blacksburg, VA, USA(弗吉尼亚理工学院计算机科学系) Department of Electrical and Computer Engineering, Virginia Tech, Blacksburg, VA, USA(弗吉尼亚理工学院电气与计算机工程系) Department of Computer Science, Carnegie Mellon University, Pittsburgh, PA, USA(卡内基梅隆大学计算机科学系) Department of Biomedical Data Science, Stanford University, Stanford, CA, USA(斯坦福大学生物医学数据科学系)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 针对现有基础模型架构直接迁移至生物学领域时忽视生物数据独特性质的问题,提出BioArc框架,利用神经架构搜索系统探索架构设计空间,发现高性能架构并提炼设计原则,同时提出架构预测方法以高效预测新任务的最优架构。

Comments Accepted at the 43nd International Conference on Machine Learning (ICML 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16060 2026-05-29 cs.LG cs.AI stat.ME stat.ML 81%

Beyond Accuracy: Are Time Series Foundation Models Well-Calibrated?

超越准确性:时间序列基础模型是否良好校准?

Coen Adler, Yuxin Chang, Felix Draxler, Samar Abdi, Padhraic Smyth

机构 * Department of Computer Science(计算机科学系) Department of Statistics(统计学系) Google, Irvine(谷歌(伊文斯堡))

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 本文系统评估了五个时间序列基础模型和两个基线的校准特性,发现基础模型校准优于基线且无系统性过度自信或信心不足。

Comments Published as a conference paper at ICLR 2026

Journal ref Proceedings of ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27178 2026-05-27 cs.CV cs.AI cs.LG cs.RO 81%

FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentation

FoundObj: 自监督基础模型作为无标签3D物体分割的奖励

Zihui Zhang, Zhixuan Sun, Yafei Yang, Jinxi Li, Jiahao Chen, Bo Yang

机构 * Shenzhen Research Institute, The Hong Kong Polytechnic University(深圳研究院,香港理工大学) vLAR Group, The Hong Kong Polytechnic University(vLAR小组,香港理工大学)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 提出FoundObj框架,利用自监督2D/3D基础模型的语义和几何先验作为奖励,通过强化学习引导超点合并,实现无标注复杂场景3D物体分割。

Comments ICML 2026. Zihui and Zhixuan are co-first authors. Code and data are available at: https://github.com/vLAR-group/FoundObj

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.19470 2026-05-20 cs.CL cs.LG 81%

Drifting Objectives for Refining Discrete Diffusion Language Models

漂移目标用于细化离散扩散语言模型

Daisuke Oba, Hiroki Furuta, Naoaki Okazaki

机构 * Institute of Science Tokyo(东京科学研究院) AIST(日本产业技术综合研究所) NII LLMC(日本信息处理学会LLMC)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

AI总结 本文研究如何将漂移方法应用于离散扩散语言模型,通过引入TokenDrift目标,将类别预测提升为软令牌特征,并在冻结语义空间中应用反称漂移,从而提升生成质量。

Comments Project page: https://daioba.github.io/tokendrift/

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.16516 2026-05-19 cs.HC cs.AI cs.CL cs.CY 81%

Alignment Drift in Long-Term Human-LLM Interaction: A Mechanism-Oriented Framework

长期人类-大语言模型交互中的对齐漂移:一种机制导向的框架

Xintong Yao

机构 * Xintong Yao(姚新同)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出一种机制导向的框架,用于描述长期人类-大语言模型交互中的对齐漂移现象,通过反馈回路和子模式选择解释漂移的发展过程,并将对齐漂移视为递归互动过程而非孤立模型失败。

Comments 16 pages, 1 appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.12522 2026-05-14 cs.CL cs.AI 81%

Differences in Text Generated by Diffusion and Autoregressive Language Models

扩散模型与自回归语言模型生成文本的差异

Zeyang Zhang, Chengwei Liang, Xingyan Chen, Meiqi Gu, Minrui Luo, Jingzhao Zhang, Tianxing He

机构 * Shanghai Qi Zhi Institute(上海启智研究院) Institute for Interdisciplinary Information Sciences, Tsinghua University(清华大学交叉信息研究院) Xiongan AI Institute(雄安人工智能研究院)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 研究比较了扩散模型与自回归语言模型在生成文本中的差异,发现扩散模型在语义连贯性和多样性上表现更优,但熵较低,这主要归因于双向上下文和解码算法的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.10640 2026-05-12 cs.CL cs.AI 81%

Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm

朝着理解语言模型持续事实知识获取的理解:从理论到算法

Haoyu Wang, Yifan Shang, Zhongxiang Sun, Weijie Yu, Xiao Zhang, Jun Xu

机构 * Gaoling School of Artificial Intelligence, Renmin University of China, Beijing, China(中国人民大学北京校区人工智能学院) School of Artificial Intelligence(人工智能学院) Data Science, University of International Business(国际商务大学数据科学)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出理论框架解释持续事实知识获取机制,提出STOC方法提升知识持续性,实验验证其有效性。

Comments Accepted by ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12362 2026-05-11 cs.LG cs.AI 81%

HYPER: A Foundation Model for Inductive Link Prediction with Knowledge Hypergraphs

HYPER:一种用于知识超图归纳链接预测的基础模型

Xingyue Huang, Mikhail Galkin, Michael M. Bronstein, İsmail İlkan Ceylan

机构 * University of Oxford(牛津大学) Google Research(谷歌研究) AITHYRA TU Wien(维也纳技术大学)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 HYPER是一种基础模型,能处理包含新实体和新关系的知识超图归纳链接预测,通过编码超边中的实体及其位置实现跨不同关系类型的迁移学习。

详情

展开后加载摘要…

URL PDF HTML 收藏