arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12169 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12169 篇

2209.14927 2023-02-27 cs.CV cs.HC cs.LG 83%

Spotlight: Mobile UI Understanding using Vision-Language Models with a Focus

Gang Li, Yang Li

专题命中 其他LLM :language model(title,abstract);prompting(abstract);分类 cs.LG

Comments Published as a conference paper at ICLR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.10461 2022-12-21 cs.CL 83%

Go-tuning: Improving Zero-shot Learning Abilities of Smaller Language Models

Jingjing Xu, Qingxiu Dong, Hongyi Liu, Lei Li

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.03299 2022-11-17 cs.CL 83%

Atlas: Few-shot Learning with Retrieval Augmented Language Models

Gautier Izacard, Patrick Lewis, Maria Lomeli, Lucas Hosseini, Fabio Petroni, Timo Schick, Jane Dwivedi-Yu, Armand Joulin, Sebastian Riedel, Edouard Grave

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01335 2022-06-14 cs.SE cs.LG 83%

Code Generation Tools (Almost) for Free? A Study of Few-Shot, Pre-Trained Language Models on Code

Patrick Bareiß, Beatriz Souza, Marcelo d'Amorim, Michael Pradel

专题命中 其他LLM :language model(title,abstract);prompting(abstract);分类 cs.LG

Comments 12 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.11070 2021-09-16 cs.CL 83%

Attribute Alignment: Controlling Text Generation from Pre-trained Language Models

Dian Yu, Zhou Yu, Kenji Sagae

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

Journal ref EMNLP 2021 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.11390 2021-04-26 cs.CL 83%

Transfer training from smaller language model

Han Zhang

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments 7 pages 7 figures. arXiv admin note: text overlap with arXiv:2103.14636 by other authors

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.00577 2020-07-30 cs.LG cs.PL stat.ML 83%

Structural Language Models of Code

Uri Alon, Roy Sadaka, Omer Levy, Eran Yahav

专题命中 其他LLM :language model(title,abstract);SLM(abstract);分类 cs.LG

Comments Appeared in ICML'2020

详情

展开后加载摘要…

URL PDF HTML 收藏
cs/0110015 2009-11-30 cs.CL 83%

Richer Syntactic Dependencies for Structured Language Modeling

Ciprian Chelba, Peng Xu

专题命中 其他LLM :language model(title,abstract);SLM(abstract);分类 cs.CL

Comments Proceedings of ASRU 2001, 4 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07717 2024-08-09 cs.RO 83%

Reflectance Estimation for Proximity Sensing by Vision-Language Models: Utilizing Distributional Semantics for Low-Level Cognition in Robotics

Masashi Osada, Gustavo A. Garcia Ricardez, Yosuke Suzuki, Tadahiro Taniguchi

专题命中 其他LLM :language model(title,abstract);large language model(abstract);foundation model(comments)

Comments 24 pages, 13 figures, submitted to Advanced Robotics Special Issue on Real-World Robot Applications of the Foundation Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.13259 2024-03-21 cs.SE 83%

Creative and Correct: Requesting Diverse Code Solutions from AI Foundation Models

Scott Blyth, Markus Wagner, Christoph Treude

专题命中 其他LLM :foundation model(title,abstract);prompting(abstract)

Comments 4 pages,Forge 2024

Journal ref AI Foundation Models and Software Engineering (FORGE '24), April 14, 2024, Lisbon, Portugal

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08195 2026-08-20 physics.comp-ph cs.AI cs.LG cs.MA physics.chem-ph 82%

Aitomia: Your Intelligent Assistant for AI-Driven Atomistic and Quantum Chemical Simulations

Aitomia:用于AI驱动的原子和量子化学模拟的智能助手

Jinming Hu, Hassan Nawaz, Yi-Fan Hou, Yuting Rui, Lijie Chi, Yuxinxin Chen, Arif Ullah, Pavlo O. Dral

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI、cs.LG

AI总结 Aitomia结合LLM和MLatom平台,通过Gaussian、ORCA等接口支持AI驱动的原子模拟和传统量子化学计算,降低原子模拟门槛,推动相关领域研究。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06114 2026-08-20 cs.CL cs.AI 版本更新 82%

Making Implicit Premises Explicit in Logical Understanding of Enthymemes

在演绎推理中显式化隐含前提

Xuyao Feng, Anthony Hunter

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种整合大型语言模型和神经符号推理器的管道,用于将隐含前提显式化并解码逻辑蕴含。

Comments Accepted at the 17th International Conference on Scalable Uncertainty Management (SUM 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.17956 2026-08-19 cs.LG cs.AI cs.SY eess.SY 新提交 82%

An Omitted Mode Is a Rare Rule: The Sampling-Verification Danger Law in Continuous Code World Models

被遗漏的模式即罕见规则:连续代码世界模型中的采样-验证危险定律

Javier Aguilar Martín

机构 * AGILabs

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI、cs.LG

AI总结 该研究揭示连续代码世界模型中采样-验证机制的危险,通过理论分析与实验发现,LLM合成的模式盲模型易被利用,接受仅能证明样本一致性,无法保证连续控制任务的性能。

Comments 92 pages, 5 figures. Code, data and result artifacts: https://github.com/JaviMaligno/code-world-models

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.17471 2026-08-19 cs.AI cs.LG 新提交 82%

When AI Designs AI: Innovation or Imitation?

当AI设计AI:创新还是模仿?

Yikang Yang, Zhengxin Yang, Luzhou Peng, Minghao Luo, Yanqi Kan, Wanling Gao, Jianfeng Zhan

机构 * University of Chinese Academy of Sciences(中国科学院大学) Northwestern University(西北大学)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI、cs.LG

AI总结 该研究对比LLM智能体与人类设计AI方法的性能和算法差异,发现智能体偶达SOTA但无法泛化,96.8%的设计属人类推导空间,多为算法选择的复用重组。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.16909 2026-08-19 cs.CY cs.AI cs.CL 新提交 82%

When Personalization Becomes Bias: Structural and Discursive Religious Framing in AI-Generated Financial Advice

当个性化成为偏见:AI生成金融建议中的结构性与话语性宗教框架

Muhammad Salar Khan, Hamza Umer, Hasan Mahmud, Sandra Rothenberg

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 该研究通过432次模拟互动分析3种LLMs在金融建议中的宗教偏见,揭示其结构性与话语性偏见机制,提出双维度框架并指出个性化与中立性的管理困境及相关启示。

Comments 50 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.14944 2026-08-18 cs.RO cs.CL cs.LG 新提交 82%

SkillComposer: Learning Reusable Skills for Natural-Language Robot Programming

SkillComposer:学习可复用技能的自然语言机器人编程系统

John Woods, Hasti Seifi

机构 * School of Computing and Augmented Intelligence, Arizona State University(亚利桑那州立大学计算与增强智能学院)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 SkillComposer是一款仿真环境下的交互式自然语言机器人编程系统,采用生成-测试架构,通过在线库学习算法生成可复用宏技能,经实验验证可提升任务成功率与可用性并减少用户工作量。

Comments 8 pages, 6 figures. Submitted to IEEE Humanoids 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.11483 2026-08-13 cs.AI cs.LG q-bio.QM 新提交 82%

A Modular Agentic Framework for Synthetically Constrained Multi-Objective Hit-to-Lead Optimization

用于合成约束多目标先导化合物优化的模块化智能体框架

Kelvin P. Idanwekhai, Enes Kelestemur, Benjamin Strickland, Matthew Hart, Steini Davidsson, Angelos Angelopoulos, Ron Alterovitz, Marcello DeLuca, Alexander Tropsha

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI、cs.LG

AI总结 本研究提出开源模块化智能体框架SABLE,结合LLM与专用工具实现合成约束下的多目标先导化合物优化,可高效富集符合计算目标的候选集,为早期药物发现提供决策支持。

Comments 22 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26998 2026-08-12 cs.CR cs.CL cs.LG 版本更新 82%

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents

AgentSnare:学习延迟、转移和瓦解自主渗透智能体

Ruoyu Wang, Heng Zhao, Renjie Wu, Mengnan Zhao, Zhixuan Chu, Wanyu Lin, Tianhang Zheng

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 AgentSnare是一种轨迹自适应欺骗系统,通过动态构建诱饵环境,吸收渗透智能体工具调用、转移其轨迹并瓦解攻击,在CVE-Bench实验中成功阻止了所有真实目标被利用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.08383 2026-08-11 cs.CL cs.LG 新提交 82%

Safety Cost of Steering Vectors Is Separable and Reducible

引导向量的安全代价是可分离且可降低的

Yuxiao Li, Gjergji Kasneci

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL、cs.LG

AI总结 本研究针对引导向量破坏LLM安全机制的问题,通过识别并移除其安全退化分量,提出带约束的原始对偶优化方法,可在保留引导效果的同时大幅降低安全退化。

Comments COLM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.06750 2026-08-10 cs.CL cs.AI 新提交 82%

Progressive Content Refinement with Decaying Reward Joint LinUCB

结合衰减奖励的渐进式内容优化与联合LinUCB

Shion Ishikawa, Pablo Loyola, Young-joo Chung, Yun Ching Liu

机构 * Rakuten Group, Inc.(乐天集团)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 针对现有迭代优化方法忽略奖励衰减导致过度利用的问题,提出结合奖励衰减建模与EM算法的联合LinUCB算法,在Sentiment Reversal和GSM8K基准上显著优于强基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.27614 2026-07-31 cs.CL cs.AI 新提交 82%

DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation

DualAnchor:无注释手语翻译中保留语言先验并提升词汇保真度

Hongbin Zhang, Junhao Liu, Xuefeng Bai, Youcheng Pan, Yang Xiang, Kehai Chen

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 针对现有无注释手语翻译方法的语言先验退化与词汇保真度差距问题,提出结合TPA和OTA的DualAnchor框架,在两个基准数据集上取得良好性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.05000 2026-07-21 cs.LG cs.AI 版本更新 82%

Automated Reinforcement Learning: An Overview

自动化强化学习:综述

Reza Refaei Afshar, Joaquin Vanschoren, Uzay Kaymak, Rui Zhang, Yaoxin Wu, Wen Song, Yingqian Zhang

机构 * Eindhoven University of Technology(埃因霍温理工大学) Jheronimus Academy of Data Science(赫尔蒙尼乌斯数据科学学院) Shandong University(山东大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文综述了自动化强化学习的发展,包括基于大语言模型的技术,并探讨了其挑战和未来研究方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.16398 2026-07-01 cs.CY cs.CL cs.LG 版本更新 82%

White-Box Sensitivity Auditing with Steering Vectors

白盒敏感性审计与引导向量

Hannah Cyberey, Yangfeng Ji, David Evans

机构 * University of Virginia(弗吉尼亚大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出白盒敏感性审计框架,通过激活引导进行更严格的模型内部评估,用于检测大语言模型中的偏见,揭示模型对保护属性的依赖。

Comments Accepted to Transactions on Machine Learning Research (TMLR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26383 2026-06-26 cs.LG cs.AI cs.AR cs.MA cs.PF 新提交 82%

SOLAR: AI-Powered Speed-of-Light Performance Analysis

SOLAR: AI驱动的光速性能分析

Qijing Huang, Sana Damani, Zhifan Ye, Athinagoras Skiadopoulos, Siva Kumar Sastry Hari, Jason Clemons, Sahil Modi, Jingquan Wang, Aditya Kane, Edward C Lin, Humphrey Shi, Christos Kozyrakis

机构 * NVIDIA(英伟达)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI、cs.LG

AI总结 提出SOLAR框架,自动从PyTorch和JAX代码推导理论最小执行时间(光速界限),通过LLM前端和确定性分析实现无违反界限的多保真度性能分析。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.18309 2026-06-18 cs.LG cs.AI 新提交 82%

SAGE: Retain-Aware Post-Hoc Sanitization of Final Unlearning Vector

SAGE: 保留感知的最终遗忘向量事后净化

Jingyuan Zhang, Yucheng Bai, Peixi Wen, Zhehao Huang, Zhengbao He, Hanling Tian, Xinwen Cheng, Haiyin Ran, Xiaolin Huang

机构 * Institute of Image Processing and Pattern Recognition, Shanghai Jiao Tong University(上海交通大学图像处理与模式识别研究所)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 提出SAGE方法,通过事后净化最终更新向量,在不重新运行原始遗忘流程的情况下,缓解大语言模型遗忘与保留能力之间的权衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15994 2026-06-16 cs.AI cs.LG 新提交 82%

Agentic Framework for Deep Learning workload migration via In-Context Learning

基于上下文学习的深度学习工作负载迁移智能体框架

Qiyue Liang, Steven Ingram, George Vanica, Andi Gavrilescu, Newfel Harrat, Hassan Sipra, Sethuraman Sankaran

机构 * Google(谷歌)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 提出结合上下文学习与Oracle驱动的自调试的自主系统,实现从PyTorch到JAX的深度学习模型自动迁移,在神经模块上达到91%数值等价性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01656 2026-06-12 cs.CL cs.AI cs.CY cs.HC physics.soc-ph 版本更新 82%

Authorship Attribution in Multilingual Machine-Generated Texts

多语言机器生成文本的作者归属

Lucio La Cava, Dominik Macko, Róbert Móro, Ivan Srba, Andrea Tagarelli

机构 * DIMES Department, University of Calabria(卡利博大学DIMES系) Kempelen Institute of Intelligent Technologies(智能技术研究所)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 提出多语言作者归属问题,研究单语言方法在18种语言和8个生成器上的跨语言迁移能力,发现显著局限。

Comments Accepted at ACL 2026 - Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.11371 2026-06-11 cs.CL cs.AI eess.AS eess.SP 新提交 82%

The Dynamics of Human and AI-Generated Language: How Semantics Fluctuates across Different Timescales

人类与AI生成语言的动态:语义如何在不同时间尺度上波动

Han-Jen Chang, Yasir Çatal, Angelika Wolman, Agustín Ibáñez, David Smith, I-Wen Su, Kai-Yuan Cheng, Georg Northoff

机构 * University of California, Berkeley(加州大学伯克利分校)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 提出语义时间尺度分析流程,通过自相关窗口度量(ACW-0)量化人类与AI生成语音中语义特异性与上下文相似性的时间组织,发现ACW-0长度与词汇通用性相关,且该关联在随机化后被削弱。

Comments 45 pages, 4 figures, 4 tables. Accepted manuscript; published in Computer Speech & Language

Journal ref Computer Speech & Language (2026) 102013

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.08919 2026-06-09 cs.AI cs.CR cs.LG 新提交 82%

Oversight Has a Capacity: Calibrating Agent Guards to a Subjective, Fatiguing Human

监督具有容量:将智能体守卫校准到主观且易疲劳的人类

Emre Turan

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI、cs.LG

AI总结 针对LLM智能体动作审批中人类评审者主观且易疲劳的问题,提出将守卫建模为成本敏感的选择性分类,并引入负载感知策略,发现过度监督反而降低安全性,形成倒U型曲线。

Comments 12 pages, 4 figures. Code and interactive demo: https://github.com/turangenesis/headroom

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.01567 2026-06-09 cs.CR cs.AI cs.CL 版本更新 82%

Defenses & Enablers For Skill Injection Attacks on Terminal Based Agents

针对终端代理的技能注入攻击的防御与使能因素

Yoshinari Fujinuma, Varun Gangal, Traian Rebedea, Makesh Narsimhan Sreedhar, Prasoon Varshney, Rebecca Qian, Anand Kannappan

机构 * Patronus AI NVIDIA

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 研究基于大语言模型的代理在重用技能时面临的安全威胁,提出守护者防御(动态和静态)将攻击成功率降低过半,并测试了攻击重述的鲁棒性。

Comments First version, small updates and clarifications likely in v2

详情

展开后加载摘要…

URL PDF HTML 收藏